Site Reliability Engineer II

JPMorgan Chase & Co.

Bengaluru

On-site

INR 1,800,000 - 3,000,000

Full time

8 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

JPMorgan Chase & Co. in Bengaluru, India is seeking a Site Reliability Engineer II to strengthen production systems through observability, automation, and resiliency engineering.

You will collaborate with Engineering and Infrastructure teams to build scalable, self-healing platforms and ensure high reliability of payments technology services. The role emphasizes ownership of end-to-end solutions, proactive problem management, and strong partnership across functions to advance reliability-first

Qualifications

  • 3+ years of experience in troubleshooting, resolving, and maintaining information technology services.
  • Experience with observability, monitoring, and incident management is required.
  • Ability to code in at least one programming language.

Responsibilities

  • Execute small to medium-sized projects independently and progressively take ownership of designing and delivering solutions end-to-end.
  • Leverage engineering best practices to develop scalable, maintainable, and resilient solutions that improve operational stability and efficiency.
  • Analyze, troubleshoot, and resolve production incidents, driving root-cause identification and permanent corrective actions.
  • Improve platform reliability, availability, and operational stability through proactive problem management and resiliency initiatives.
  • Design, implement, and enhance observability capabilities, including monitoring, alerting, dashboards, telemetry, SLIs, and SLOs.
  • Monitor production environments, identify anomalies and trends, and proactively address risks using standard observability and operational tooling.
  • Eliminate operational toil through automation, self-healing solutions, process optimization, and reuse-first engineering practices.
  • Support incident, problem, and change management processes across applications, infrastructure, and full-stack technology services.
  • Partner with Engineering, Infrastructure, Product, and Operations teams to influence design decisions with a reliability-first mindset.
  • Utilize enterprise-approved AI and agentic capabilities to accelerate incident triage, root-cause analysis, and remediation while adhering to data security and governance standards.
  • Continuously improve service resilience, operational efficiency, and customer experience by driving automation, observability, and reliability engineering best practices.

Skills

Observability
Automation
Incident management
Programming
Kubernetes
Containers
CI/CD
Networking
Collaboration

Tools

Grafana
Dynatrace
Prometheus
Datadog
Splunk
Jenkins
GitLab
Terraform
Kubernetes

Job description

Join a dynamic team shaping the tech backbone of our operations, where your expertise fuels seamless system functionality and innovation.

As a Site Reliability Engineer II at JPMorgan Chase within the Commercial & Investment Bank- Payments Technology team, you will use technology to solve business problems and leverage software engineering best practices as we strive towards excellence. The Site Reliability Engineer is responsible for ensuring the reliability, availability, performance, and operational excellence of production services through observability, automation, incident management, resiliency engineering, and continuous improvement. The role partners closely with Engineering and Infrastructure teams to build scalable, self-healing, and highly resilient platforms.

Job responsibilities
  • - Execute small to medium-sized projects independently and progressively take ownership of designing and delivering solutions end-to-end.
  • - Leverage engineering best practices to develop scalable, maintainable, and resilient solutions that improve operational stability and efficiency.
  • - Analyze, troubleshoot, and resolve production incidents, driving root-cause identification and permanent corrective actions.
  • - Improve platform reliability, availability, and operational stability through proactive problem management and resiliency initiatives.
  • - Design, implement, and enhance observability capabilities, including monitoring, alerting, dashboards, telemetry, SLIs, and SLOs.
  • - Monitor production environments, identify anomalies and trends, and proactively address risks using standard observability and operational tooling.
  • - Eliminate operational toil through automation, self-healing solutions, process optimization, and reuse-first engineering practices.
  • - Support incident, problem, and change management processes across applications, infrastructure, and full-stack technology services.
  • - Partner with Engineering, Infrastructure, Product, and Operations teams to influence design decisions with a reliability-first mindset.
  • - Utilize enterprise-approved AI and agentic capabilities to accelerate incident triage, root-cause analysis, and remediation while adhering to data security and governance standards.
  • - Continuously improve service resilience, operational efficiency, and customer experience by driving automation, observability, and reliability engineering best practices.
Required qualifications, capabilities, and skills
  • - 3+ years of experience or equivalent expertise troubleshooting, resolving, and maintaining information technology services
  • - Ability to code in at least one programming language
  • - Familiar with site reliability concepts, principles, and practices
  • - Familiar with observability such as white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, and others
  • - Familiarity with containers or a common Server OS such as Linux and Windows
  • - Emerging knowledge of software, applications and technical processes within a given technical discipline (e.g., Cloud, artificial intelligence, Android, etc.)
  • - Emerging knowledge of continuous integration and continuous delivery tools like Jenkins, GitLab, or Terraform
  • - Emerging knowledge of common networking technologies
  • - Ability to work in a large, collaborative team and demonstrates the willingness to vocalize ideas with peers and managers
  • - Understanding of how to prioritize and adjust work plans to adapt to changes in assigned responsibilities and projects
  • - Eagerness to participate in learning opportunities to enhance one’s effectiveness in executing day-to-day project activities
  • - Ability to demonstrate and apply existing and new system processes, methodologies, and skills to contribute to the development of systems
  • - Strong partnership skills with understand of escalation management across partner relationship
  • - Knowledge of applications or infrastructure in a large-scale technology environment on premises or public cloud
Preferred qualifications, capabilities, and skills
  • - Knowledge of one or more general purpose programming languages or automation scripting
  • - Desire to grow and learn in the AI space as the business grows
  • - Knowledge of Kubernetes, ITRS Active Console, Splunk, and Dynatrace
  • - Knowledge of the Payments Technology
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II
Site Reliability Engineer II

JPMorganChase • Bengaluru

On-site
INR 2,000,000 - 3,200,000
Site Reliability Engineer
Site Reliability Engineer

JP Morgan Services India Pvt Ltd • Bengaluru

On-site
INR 1,500,000 - 2,300,000
Site Reliability Engineer II
Site Reliability Engineer II

Next Frontier Capital • Bengaluru

On-site
INR 4,500,000 - 6,500,000
Site Reliability Engineer II - Java/Python, Kubernetes, AWS, Terraform
Site Reliability Engineer II - Java/Python, Kubernetes, AWS, Terraform

JPMorgan Chase & Co. • Bengaluru

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineer II - Java/Python, Kubernetes, AWS, Terraform
Site Reliability Engineer II - Java/Python, Kubernetes, AWS, Terraform

JPMorganChase • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

JP Morgan Services India Pvt Ltd • Bengaluru

On-site
INR 3,000,000 - 4,200,000
Software Engineer II - Python And AWS
Software Engineer II - Python And AWS

JPMorganChase • Hyderabad

On-site
INR 2,500,000 - 4,000,000
Site Reliability Engineer III
Site Reliability Engineer III

JPMorgan Chase & Co. • Hyderabad

On-site
INR 1,500,000 - 2,000,000
Software Engineer II - Python & AWS
Software Engineer II - Python & AWS

JPMorganChase • Hyderabad

On-site
INR 1,400,000 - 2,100,000
Site Reliability Engineer III
Site Reliability Engineer III

JPMorganChase • Mumbai

On-site
INR 4,000,000 - 7,000,000