SRE Software Engineer III

JPMorgan Chase & Co.

Kentucky

On-site

USD 120,000 - 180,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

JPMorgan Chase & Co. within Consumer & Community Banking seeks a Site Reliability Engineer (Software Engineer III) to balance hands-on development with reliability at scale.

You will design, build, and run resilient services, improve production stability through automation, and work with product teams to enable safe delivery. You will lead incidents, drive post-incident reviews, and implement observability standards while leveraging enterprise AI-assisted tooling to boost code quality and

Qualifications

  • Formal training or certification in software engineering concepts and 3+ years of experience.
  • Experience developing production-grade systems in Python, Java, or similar languages.
  • Experience operating and improving reliability of distributed systems on Linux/Unix.
  • Cloud and virtualization concepts, APIs, and scalable services.
  • Observability and diagnostics using monitoring/logging platforms.
  • Version control and CI/CD with secure engineering practices.
  • Strong communication and incident leadership across teams.
  • Experience with enterprise AI-assisted development tools and secure usage.

Responsibilities

  • Engineer and improve reliability for large-scale, distributed services.
  • Design automation to reduce manual work and improve recovery time.
  • Lead incident response and blameless post-incident reviews.
  • Collaborate with development teams to enable safe, repeatable deployments.
  • Implement self-healing patterns and capacity management to boost availability.
  • Develop end-to-end observability with metrics, logs, and traces.
  • Analyze incidents to anticipate and prevent customer-impacting issues.
  • Participate in on-call rotations and software delivery responsibilities.
  • Apply AI-assisted development tools with peer review and secure coding standards.

Skills

Python/Java
Linux/Unix
Cloud computing
CI/CD
Observability
Incident leadership
Security best practices
AI-assisted development
Effective communication
Git workflows

Education

Software engineering certification

Tools

Kubernetes
Dynatrace
Splunk
Grafana
Git
Tomcat
Nginx
Oracle
MySQL

Job description

You will balance hands-on development with technical resiliency.

As a Site Reliability Engineer (Software Engineer III) at JPMorganChase within Consumer & Community Banking, you will help build and run resilient, scalable services by combining software engineering with operational excellence. You will improve production stability through automation, reliability engineering, and disciplined incident management, partnering across engineering and product teams to reduce risk and accelerate safe delivery.

Job Responsibilities
  • Engineer and improve reliability for large-scale, distributed services by applying software and systems engineering best practices to production operations.
  • Design and deliver automation that reduces manual operational work, improves recovery time, and strengthens production stability and service resilience.
  • Lead response and coordination for high-priority incidents, drive issue triage to resolution, and facilitate blameless post-incident reviews that result in measurable fixes.
  • Partner with application development teams throughout the software delivery lifecycle to enable sustainable releases, safer change practices, and repeatable deployments.
  • Implement self-healing patterns, resilience strategies, and capacity management practices to improve availability and reduce operational toil.
  • Build and evolve end-to-end observability (metrics, logs, traces) and alerting standards to enable actionable monitoring and reduce noise.
  • Apply data-driven analysis of incidents and usage patterns to anticipate reliability risks and proactively prevent customer-impacting issues.
  • Support a balanced operating model that includes both engineering delivery and operational responsibilities, including participation in on-call rotations as needed.
  • Leverages enterprise-authorized AI coding assist tools within the work environment to improve code quality, delivery speed, and productivity across complex deliverables (e.g., code generation/refactoring, unit test creation, documentation), while validating outputs through peer review, automated testing, and secure coding standards; contributes learnings and reusable patterns to improve broader team effectiveness.
  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation.
Required Qualifications, Capabilities, and Skills
  • Formal training or certification on software engineering concepts and 3+ years applied experience
  • Hands-on software development experience in one or more general-purpose languages (e.g., Python, Java, shell scripting) supporting production-grade systems.
  • Experience operating and improving reliability of distributed systems on Linux/Unix environments, including troubleshooting across application, infrastructure, and data layers.
  • Experience with cloud and virtualization concepts, APIs, and modern engineering practices for scalable and fault-tolerant services.
  • Experience building observability and operational diagnostics using monitoring/logging platforms (e.g., Dynatrace, Splunk, Grafana, cloud-native telemetry tools).
  • Working knowledge of version control and modern delivery practices (e.g., Git-based workflows, continuous integration/continuous delivery) with a focus on quality and secure engineering.
  • Strong critical thinking, incident leadership, and communication skills, with the ability to partner effectively across engineering, product, and operations stakeholders.
  • Hands-on experience using enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, test creation, troubleshooting, or documentation) with demonstrated ability to critically evaluate, validate, and refine AI-generated outputs for correctness, performance, and security.
  • Understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; ability to guide peers on safe and effective usage within team practices.
Preferred Qualifications, Capabilities, and Skills
  • Experience with Kubernetes and containerized workloads, including deployment patterns and operational troubleshooting.
  • Experience designing frameworks that improve developer experience, release velocity, code health, and engineering standards.
  • Experience with performance testing, bottleneck analysis, and capacity planning for high-throughput services.
  • Experience administering application servers, web servers, and databases (e.g., Tomcat, Nginx, Oracle, MySQL) in production contexts.
  • Certifications in cloud architecture or data platforms (e.g., AWS Solutions Architect or equivalent).
  • 5+ years of related industry experience across application development, site reliability engineering, or DevOps in large-scale environments.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

DevOps Software Engineer III
DevOps Software Engineer III

JPMorgan Chase & Co. • Westerville (OH)

On-site
USD 110,000 - 160,000
DevOps Software Engineer III
DevOps Software Engineer III

JPMorgan Chase & Co. • Columbus (OH)

On-site
USD 110,000 - 150,000
Software Engineer III- SRE
Software Engineer III- SRE

JPMorgan Chase & Co. • Wilmington (DE)

On-site
USD 110,000 - 150,000
Site Reliability Engineer III
Site Reliability Engineer III

JPMorgan Chase & Co. • Irvine (CA)

On-site
USD 130,000 - 170,000
Site Reliability Engineer III
Site Reliability Engineer III

JPMorgan Chase & Co. • Houston (TX)

On-site
USD 120,000 - 180,000
Site Reliability Engineer II
Site Reliability Engineer II

Talentify • Chicago (IL)

On-site
USD 120,000 - 180,000
Site Reliability Engineer III - Machine Learning
Site Reliability Engineer III - Machine Learning

JPMorgan Chase & Co. • Wilmington (DE)

On-site
USD 120,000 - 160,000
Site Reliability Engineer III
Site Reliability Engineer III

JPMorgan Chase & Co. • Chicago (IL)

On-site
USD 120,000 - 180,000
Restaurant that supports hybrid work
Site Reliability Engineer III- Production Management
Site Reliability Engineer III- Production Management

JPMorgan Chase & Co. • City of Rochester (NY)

On-site
USD 120,000 - 160,000
Site Reliability Engineer III- Production Management
Site Reliability Engineer III- Production Management

JPMorgan Chase & Co. • New York (NY)

On-site
USD 140,000 - 190,000