Site Reliability Engineer II

American Express

Chennai District

On-site

INR 1,800,000 - 3,200,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive base salary
Bonus incentives
Health and wellness benefits
Hybrid/onsite work options
Career development opportunities

Job summary

American Express seeks a Site Reliability Engineer II in Chennai to collaborate with software and tech teams to improve resilience, scalability and performance through feature work, automation, and architectural input.

You will work with Kubernetes/Docker in cloud environments (AWS/Azure/GCP), implement chaos testing, and contribute to disaster recovery planning and SRE best practices with support from senior colleagues.

Qualifications

  • Bachelor’s degree in Computer Science, IT, Engineering, or equivalent.
  • Knowledge of modern observability stack – Splunk, Elastic Search, Prometheus, Grafana.
  • Knowledge of containerization technologies (e.g., Kubernetes, Docker) and microservices architecture.
  • Knowledge of observability tools and methodologies, including logging, monitoring, tracing, and performance analysis platforms.
  • Knowledge of cloud‑based Site Reliability Engineering (SRE) practices and experience with public cloud platforms such as AWS, Azure, or Google Cloud.

Responsibilities

  • Collaborates with Software Engineering teams to design, develop, and implement features that enhance system resilience, scalability, and performance, while identifying and addressing potential system bottlenecks and failure points with guidance from senior colleagues
  • Develops and implements automation tools and frameworks, including infrastructure as code (IaC) practices to streamline operational workflows, deployment processes, and infrastructure management, with guidance from peers and leaders
  • Collaborates with senior engineers to contribute to the architectural design of systems, ensuring that reliability, scalability, and performance considerations are integrated into design discussions and decision‑making processes
  • Collaborates in the design and execution of chaos engineering experiments and other resiliency testing, analyzing results and implementing improvements to enhance system robustness and recovery capabilities, with guidance from peers and leaders
  • Develops and implements of disaster recovery plans and business continuity strategies, ensuring systems can recover quickly and effectively from unexpected disruptions
  • Collaborates with seniors to promote and implement best practices such as error budgeting, service-level objectives (SLOs), and service-level indicators (SLIs), contributing to a culture of continuous improvement and reliability
  • Collaborates and co‑creates effectively with teams in product and the business to align technology initiatives with business objectives

Skills

System reliability mindset
Performance tuning
Collaboration

Education

Bachelor's degree in Computer Science / IT / Engineering
Advanced degree preferred

Tools

Kubernetes
Docker
Splunk
Elastic Search
Prometheus
Grafana
AWS
Azure
GCP

Job description

Job Description

Site Reliability Engineer II collaborates with engineering teams to enhance system resilience, scalability, and performance through feature development, automation, architectural design, resiliency testing, and disaster recovery planning, while promoting best practices for continuous improvement.

Responsibilities
  • Collaborates with Software Engineering teams to design, develop, and implement features that enhance system resilience, scalability, and performance, while identifying and addressing potential system bottlenecks and failure points with guidance from senior colleagues
  • Develops and implements automation tools and frameworks, including infrastructure as code (IaC) practices to streamline operational workflows, deployment processes, and infrastructure management, with guidance from peers and leaders
  • Collaborates with senior engineers to contribute to the architectural design of systems, ensuring that reliability, scalability, and performance considerations are integrated into design discussions and decision‑making processes
  • Collaborates in the design and execution of chaos engineering experiments and other resiliency testing, analyzing results and implementing improvements to enhance system robustness and recovery capabilities, with guidance from peers and leaders
  • Develops and implements of disaster recovery plans and business continuity strategies, ensuring systems can recover quickly and effectively from unexpected disruptions
  • Collaborates with seniors to promote and implement best practices such as error budgeting, service-level objectives (SLOs), and service-level indicators (SLIs), contributing to a culture of continuous improvement and reliability
  • Collaborates and co‑creates effectively with teams in product and the business to align technology initiatives with business objectives
Qualifications

Education Qualifications:

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, and/or comparable experience; advance degree preferred
  • Knowledge of modern observability stack – Splunk, Elastic Search, Prometheus, Grafana
  • Knowledge of containerization technologies (e.g., Kubernetes, Docker) and microservices architecture
  • Knowledge of observability tools and methodologies, including experience with logging, monitoring, tracing, and performance analysis platforms
  • Knowledge of cloud‑based Site Reliability Engineering (SRE) practices and experience with public cloud platforms such as AWS, Azure, or Google Cloud
Work Experience
  • Experience in software development, or technology operations, with a focus on Site Reliability Engineering
  • Experience in Linux/Unix systems, object‑oriented programming languages (e.g., Java), scripting languages (e.g., Python, Bash), and cloud platforms (e.g., AWS, Azure, GCP)
Licenses And Certifications
  • Advanced certification in Site Reliability Engineering (SRE) or related is a plus
About Us

At American Express, our culture is built on a 175-year history of innovation, shared values and Leadership Behaviors, and an unwavering commitment to back our customers, communities, and colleagues. From delivering differentiated products to providing world‑class customer service, we operate with a strong risk mindset, ensuring we continue to uphold our brand promise of trust, security, and service.

As part of Team Amex, you’ll experience our powerful backing with comprehensive support for your holistic well‑being and many opportunities to learn new skills, develop as a leader, and grow your career. Here, your voice and ideas matter, your work makes an impact, and together, you will help us define the future of American Express.

About The Team

We back you with benefits that support your holistic well‑being so you can be and deliver your best. This means caring for you and your loved ones' physical, financial, and mental health, as well as providing the flexibility you need to thrive personally and professionally:

  • Competitive base salaries
  • Bonus incentives
  • Support for financial-well‑being and retirement
  • Comprehensive medical, dental, vision, life insurance, and disability benefits (depending on location)
  • Flexible working model with hybrid, onsite or virtual arrangements depending on role and business need
  • Generous paid parental leave policies (depending on your location)
  • Free access to global on‑site wellness centers staffed with nurses and doctors (depending on location)
  • Free and confidential counseling support through our Healthy Minds program
  • Career development and training opportunities

American Express is an equal opportunity employer and makes employment decisions without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, disability status, age, or any other status protected by law.

Offer of employment with American Express is conditioned upon the successful completion of a background verification check, subject to applicable laws and regulations.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II
Site Reliability Engineer II

American Express • Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Competitive salary
Bonus incentives
Medical benefits
+2
Senior Software Engineer II - Full Stack Web Application
Senior Software Engineer II - Full Stack Web Application

American Express • Chennai District

On-site
INR 3,500,000 - 7,000,000
Competitive salary
Bonus incentives
Healthcare benefits
+5
Senior Software Engineer II
Senior Software Engineer II

American Express • Gurgaon

Hybrid
INR 3,500,000 - 6,500,000
Competitive base salaries
Bonus incentives
Medical, dental, vision, life & 24
+2
Engineer I- Openshift Engineer
Engineer I- Openshift Engineer

American Express Services Europe Limited • Bengaluru Urban

Hybrid
INR 1,200,000 - 1,600,000
Flexible working model
Generous paid parental leave
Comprehensive medical and dental benefits
+1
Software Engineer II
Software Engineer II

American Express • Bengaluru

Hybrid
INR 1,500,000 - 2,500,000
Competitive base salaries
Bonus incentives
Flexible hybrid/onsite/virtual work
Software Engineer III
Software Engineer III

American Express • Gurugram District

Hybrid
INR 1,800,000 - 2,400,000
Competitive salary
Bonus incentives
Retirement benefits
+6
Engineer I
Engineer I

American Express • Chennai

Hybrid
INR 600,000 - 1,200,000
Competitive base salaries
Bonus incentives
Comprehensive medical, dental, vision, life insurance
+5
Senior Software Engineer II
Senior Software Engineer II

American Express • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Competitive base salaries
Bonus incentives
Flexible working model
Staff Engineer
Staff Engineer

American Express Services Europe Limited • Bengaluru Urban

Hybrid
INR 1,600,000 - 2,000,000
Comprehensive medical and dental benefits
Flexible working model
Career development opportunities
Engineer - Java Full Stack Developer
Engineer - Java Full Stack Developer

American Express Services Europe Limited • Gurugram District

Hybrid
INR 1,000,000 - 1,500,000
Comprehensive medical, dental, vision insurance
Flexible working model
Career development opportunities
+1