Senior Site Reliability Engineer — AI-Driven Reliability & Cloud Ops

JPMorganChase

Houston (TX)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

JPMorgan Chase & Co. seeks a Site Reliability Engineer III to own end-to-end reliability across complex, mission-critical systems. You will enhance availability, scalability, and performance via code and cloud infrastructure, collaborating with software engineers to implement CI/CD pipelines and robust monitoring.

The role emphasizes AI-assisted incident triage and proactive problem-solving. As part of Corporate Technology, Risk Technology, you will guide design decisions, validate AI-driven

Qualifications

  • Formal training or certification on site reliability engineering concepts and 3+ years applied experience.
  • Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platform.
  • Proficient in at least one programming language such as Python, Java/Spring Boot, and .Net.
  • Working knowledge of enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivity.
  • Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements.
  • Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., Cloud, AI, Android, etc.).
  • Experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection.

Responsibilities

  • Guides and assists others in the areas of building appropriate level designs and gaining consensus from peers where appropriate, supporting adoption of site reliability engineering best practices within your team
  • Collaborates with other software engineers and teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Implements infrastructure, configuration, and network as code for the applications and platforms in your remit
  • Collaborates with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they impact customers
  • Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to SLO outcomes.
  • Familiar with availability, reliability, scalability, and solutions in their applications and works with partners to improve these outcomes iteratively
  • Proactively recognizes road blocks and identifies improvements to solve business problems, including exploring new technologies where appropriate

Skills

Python
Java/Spring Boot
.Net
SRE concepts

Education

Formal training or certification on site reliability engineering concepts

Tools

CI/CD tooling
Container orchestration
Telemetry/Monitoring

Job description

JPMorgan Chase & Co. seeks a Site Reliability Engineer III to own end-to-end reliability across complex, mission-critical systems. You will enhance availability, scalability, and performance via code and cloud infrastructure, collaborating with software engineers to implement CI/CD pipelines and robust monitoring.

The role emphasizes AI-assisted incident triage and proactive problem-solving. As part of Corporate Technology, Risk Technology, you will guide design decisions, validate AI-driven

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer - Cloud, AI-Driven Ops
Senior Site Reliability Engineer - Cloud, AI-Driven Ops

JPMorganChase • Houston (TX)

On-site
USD 140,000 - 180,000
Senior Site Reliability Engineer — AI-Driven Cloud Resilience
Senior Site Reliability Engineer — AI-Driven Cloud Resilience

Socket.dev • Wilmington (DE)

On-site
USD 120,000 - 160,000
Senior Site Reliability Engineer: AI-Driven Cloud Ops
Senior Site Reliability Engineer: AI-Driven Cloud Ops

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 120,000 - 180,000
Senior SRE - AI-Driven Reliability & Cloud Automation
Senior SRE - AI-Driven Reliability & Cloud Automation

JPMorgan Chase & Co. • Chicago (IL)

On-site
USD 130,000 - 170,000
Senior Site Reliability Engineer — AI-Driven Reliability & Cloud
Senior Site Reliability Engineer — AI-Driven Reliability & Cloud

JPMorganChase • Chicago (IL)

On-site
USD 120,000 - 180,000
Site Reliability Engineer III — AI-Driven Reliability & Cloud
Site Reliability Engineer III — AI-Driven Reliability & Cloud

JPMorgan Chase • Chicago (IL)

On-site
USD 114,000 - 155,000
Health coverage
Retirement plan
Tuition reimbursement
+1
Site Reliability Engineer III: AI-Driven Cloud & 24x7 Ops
Site Reliability Engineer III: AI-Driven Cloud & 24x7 Ops

JPMorganChase • Irvine (CA)

On-site
USD 130,000 - 170,000
Principal Site Reliability Engineer – AI-Driven Reliability
Principal Site Reliability Engineer – AI-Driven Reliability

JPMorganChase • Jersey City (NJ)

On-site
USD 180,000 - 260,000
Competitive base salary
Discretionary incentive compensation
Health care coverage
+3
Senior Principal Site Reliability Engineer (AI-Driven)
Senior Principal Site Reliability Engineer (AI-Driven)

Socket.dev • New Jersey

On-site
USD 170,000 - 260,000
SRE III: AI-Driven Reliability & Cloud Ops
SRE III: AI-Driven Reliability & Cloud Ops

JPMorgan Chase & Co. • Wilmington (DE)

On-site
USD 120,000 - 160,000