As a Site Reliability Engineer (Azure Cloud), you will make an impact by ensuring the reliability, scalability, performance, and operational excellence of mission-critical, customer-facing applications running in a 24x7 environment. You will collaborate with development, infrastructure, and operations teams to build resilient cloud solutions, automate operational processes, strengthen observability practices, and drive continuous improvements that enhance system availability and customer experience.
In this Role, You Will
- Ensure the reliability, availability, and scalability of business-critical applications while supporting high-availability service objectives.
- Design and implement automation solutions that reduce manual effort, improve operational efficiency, and enhance system stability.
- Monitor, analyze, and optimize application and platform performance, proactively identifying and resolving issues before they impact customers.
- Lead incident response, problem management, and root cause analysis efforts, driving long-term corrective actions and resiliency improvements.
- Partner with development, infrastructure, and security teams to embed reliability, observability, and compliance practices throughout the software delivery lifecycle.
Work Model
We strive to provide flexibility wherever possible. Based on business and client requirements, this role may operate in a hybrid, remote, or office-based work model. Work arrangements will be discussed during the recruitment process and may vary based on project and client needs.
What You Need to Have to Be Considered
- Experience in Site Reliability Engineering (SRE), Application Support, Production Support, or a related role supporting enterprise applications.
- Strong Linux administration and troubleshooting skills in large-scale production environments.
- Experience supporting cloud-based environments, preferably Microsoft Azure.
- Strong scripting and automation skills using Python, Bash, or similar technologies.
- Understanding of enterprise technology environments, including networking, databases, messaging platforms, authentication, and access management concepts.
- Experience supporting microservices-based applications and distributed systems.
- Knowledge of web and internet technologies, including APIs, HTTP/HTTPS, encryption, and service communications.
- Experience participating in incident management, root cause analysis, and operational support activities.
- Strong analytical, communication, collaboration, and problem-solving skills.
- Passion for reliability, performance, scalability, automation, and continuous improvement.
- Bachelor's degree in Computer Science, Engineering, Information Technology, or equivalent professional experience.
These Will Help You Stand Out
- Experience with observability and monitoring platforms such as Dynatrace, Datadog, Splunk, Prometheus, Grafana, or similar technologies.
- Experience working in highly regulated environments, including financial services organizations.
- Knowledge of security and compliance frameworks such as PCI DSS, NIST, CIS, or similar standards.
- Experience supporting high-volume, customer-facing applications with demanding availability requirements.
- Familiarity with cloud-native architectures and modern operational practices.
- Understanding of networking, performance engineering, and security best practices.
- Experience supporting fintech, banking, payments, or financial technology platforms.
Leadership Expectations
- Demonstrate ownership and accountability for service reliability and operational outcomes.
- Drive continuous improvement initiatives focused on automation, efficiency, and resiliency.
- Collaborate effectively across engineering, infrastructure, security, and business teams.
- Contribute to operational excellence by sharing knowledge, documenting processes, and mentoring peers when appropriate.
- Act as a trusted technical resource during incidents and critical production events.
We're excited to meet people who share our mission and can contribute to creating exceptional outcomes for our clients, communities, and colleagues. If this role aligns with your experience and career aspirations, we encourage you to apply.
Additional employment information
Compensation information is accurate as of the date of this posting. Cognizant reserves the right to modify this information at any time, subject to applicable law.
Applicants may be required to attend interviews in person or by video conference. In addition, candidates may be required to present their current state or government issued ID during each interview.
Cognizant is an equal opportunity employer. Your application and candidacy will not be considered based on race, color, sex, religion, creed, sexual orientation, gender identity, national origin, disability, genetic information, pregnancy, veteran status or any other characteristic protected by federal, state or local laws.
Applicants for US based positions must be currently authorized to work for any employer in the United States. The company is unable to provide employment-based immigration sponsorship for US based positions.
If you have a disability that requires reasonable accommodation to search for a job opening or submit an application, please email [email protected] for roles based in the Americas or [email protected] for roles based in India.
About Cognizant:
Cognizant (Nasdaq: CTSH) is an AI Builder and technology services provider, bridging the gap between AI investment and enterprise value by building full-stack AI solutions for our clients. Our deep industry, process and engineering expertise enables us to build an organization’s unique context into technology systems that amplify human potential, drive tangible outcomes and keep global enterprises ahead in a fast-changing world. See how at cognizant.ai or @cognizant.











