Site Reliability Engineer II

JPMorgan Chase & Co.

Bengaluru

On-site

INR 1,000,000 - 1,800,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

JPMorgan Chase & Co. in Bengaluru seeks an experienced Site Reliability Engineer II to strengthen reliability across enterprise platforms. You will automate operations, support incident management, and collaborate across teams to improve services.

The role emphasizes Python/Shell scripting, Ansible, and networking expertise. The position involves working with AI-powered tools to streamline triage and post-incident analysis while ensuring security and compliance.

Qualifications

  • Formal training or certification on Site Reliability concepts and 2+ years applied experience.
  • Experience or strong exposure to network operations/engineering and incident/problem management practices.
  • Hands-on automation skills in Python and/or Shell, and Ansible for configuration/operational tasks.
  • Practical troubleshooting across routing/switching and at least one of: firewall, load balancer, proxy.
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits.
  • Ability to assess AI-assisted operational recommendations for correctness and risk, and apply appropriate controls.

Responsibilities

  • Contribute to problem management activities: evidence collection, timeline building, RCA and action items.
  • Build and enhance automation using Python, Shell scripting, and Ansible for health checks and data collection.
  • Utilize AI capabilities to speed up incident triage, troubleshooting, and post-incident analysis with proper data handling.
  • Participate in incident management for network services: monitoring, triage, troubleshooting, mitigation support, escalation.
  • Identify toil and work towards elimination via systems engineering or code updates.
  • Support and troubleshoot core networking domains: Routing, Switching, Firewalls, Load Balancers, Proxies, SD-WAN.
  • Apply AI capabilities to identify recurring toil and reliability risks with measurable SLOs.
  • Adopt SRE best practices and demonstrate reliability, toil reduction, and operational readiness.

Skills

Python scripting
Shell scripting
Ansible
Networking basics
Cloud infrastructure
Incident management
SRE concepts

Education

SRE Certification

Tools

Linux
Observability tools
CI/CD tooling

Job description

Play a key role in ensuring system reliability at one of the world's most iconic and largest financial institutions.

As a Site Reliability Engineer II at JPMorgan Chase within the Enterprise Technology - Infrastructure Platforms team, you will use technology to solve business problems and leverage software engineering best practices as we strive towards excellence. This role often works independently to execute small to medium projects, but you'll also have the opportunity to collaborate with cross functional teams to continually improve your level of knowledge about JPMorgan Chase’s business and relevant technologies.

Job responsibilities
  • Contribute to problem management activities: evidence collection, timeline building, contributing to RCA and action items
  • Build and enhance automation using Python, Shell scripting, and Ansible (e.g., basic health checks, data collection scripts, configuration validation
  • Uses enterprise-authorized AI capabilities within the work environment to speed up incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Participate in incident management for network services: monitoring, triage, troubleshooting, mitigation support, and escalation.
  • Recognizes toil within the role and proactively works towards eliminating it through systems engineering or updating application code
  • Support and troubleshoot core networking domains: Routing and Switching, Firewalls, Load Balancers, Proxies and SD-WAN, SDA, and broader software-defined networking (SND) concepts
  • Applies enterprise-authorized AI capabilities within the work environment to identify recurring toil and reliability risks from operational signals, prioritizing reuse-first improvements and measurable SLO outcomes.
  • Demonstrate SRE mindset: learn and apply concepts such as reliability, toil reduction, and operational readiness; understand NFRs and introductory FMEA concepts
  • Supports the adoption of site reliability engineering best practices within your team
  • Should complete SRE Bar Raiser Program
Required qualifications, capabilities, and skills
  • Formal training or certification on Site Reliability concepts and 2+ years applied experience
  • Experience or strong exposure to network operations/engineering and incident/problem management practices.
  • Hands-on automation skills in Python and/or Shell, and Ansible for configuration/operational tasks.
  • Practical troubleshooting across routing/switching and at least one of: firewall, load balancer, proxy.
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows (e.g., troubleshooting support and runbook drafting) with strong validation habits and awareness of data sensitivity.
  • Ability to assess AI-assisted operational recommendations for correctness and risk, and apply appropriate controls to maintain resiliency, security, and auditability.
  • Experience maintaining a cloud-based infrastructure
  • Familiar with site reliability concepts, principles, and practices
  • Familiar with observability such as white and black box monitoring, service level objective alerting, and telemetry collection
  • Familiarity with containers or a common server OS such as Linux and Windows
  • Emerging knowledge of continuous integration and continuous delivery practices and related tooling
Preferred qualifications, capabilities, and skills
  • Exposure to Cisco ACI / Fabrics.
  • Certifications such as CCNA and/or other vendor certifications.
  • Experience in a financial institution environment
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II - Python, Observability, AWS, Terraform
Site Reliability Engineer II - Python, Observability, AWS, Terraform

JPMorganChase • Mumbai

On-site
INR 1,800,000 - 2,400,000
Lead Site Reliability Engineer - Network Infrastructure
Lead Site Reliability Engineer - Network Infrastructure

JP Morgan Services India Pvt Ltd • Bengaluru

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineer III
Site Reliability Engineer III

JPMorgan Chase & Co. • Hyderabad

On-site
INR 1,500,000 - 2,000,000
Lead Infrastructure Engineer -Network Operation SME(Subject Matter Expert)
Lead Infrastructure Engineer -Network Operation SME(Subject Matter Expert)

Next Frontier Capital • Hyderabad

On-site
INR 600,000 - 900,000
Lead Infrastructure Engineer -Network Operation SME(Subject Matter Expert)
Lead Infrastructure Engineer -Network Operation SME(Subject Matter Expert)

JPMorgan Chase & Co. • Hyderabad

On-site
INR 4,200,000 - 5,400,000
Lead Infrastructure Engineer –Network Operation SME(Subject Matter Expert)
Lead Infrastructure Engineer –Network Operation SME(Subject Matter Expert)

JPMorgan Chase & Co. • Hyderabad

On-site
INR 2,500,000 - 4,200,000
Infrastructure Engineer III - Network Engineer
Infrastructure Engineer III - Network Engineer

JPMorgan Chase & Co. • Hyderabad

On-site
INR 2,000,000 - 3,500,000
None
Lead Software Engineer - DevOps/SRE/AWS/EKS
Lead Software Engineer - DevOps/SRE/AWS/EKS

JPMorganChase • Bengaluru

On-site
INR 1,800,000 - 3,200,000
Lead Infrastructure Engineer-Network
Lead Infrastructure Engineer-Network

JPMorgan Chase & Co. • Bengaluru

On-site
INR 3,000,000 - 5,000,000
Software Engineer Iii - Sre, Python/Java, Aws, Ai
Software Engineer Iii - Sre, Python/Java, Aws, Ai

JPMorganChase • Telangana

On-site
INR 1,800,000 - 2,800,000