Senior Site Reliability Engineer | Incident Commander

U.S. Bank

Chicago (IL)

On-site

USD 112,000 - 131,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Healthcare
Life insurance
Disability
Parental leave
401(k)
Paid vacation
Holidays
Adoption assistance
Sick leave

Job summary

U.S. Bank in Chicago seeks a senior Site Reliability Engineer to lead complex production incidents, perform RCA, and drive reliability across cloud platforms. You will own the incident command during major outages and collaborate with software, infra, and product teams to implement lasting fixes.

Hands-on with AWS/Azure, Kubernetes, Docker, CI/CD, and observability tools is required. You will mentor engineers, optimize monitoring and runbooks, and use MTTR/MTTD metrics to push continuous

Qualifications

  • Bachelor's degree or equivalent work experience.
  • Six to eight years of relevant work experience in IT service management, production support, or related areas.
  • Strong expertise in SRE, DevOps, incident response, and distributed systems operations.

Responsibilities

  • Lead troubleshooting and resolution of complex production incidents, including outages and performance issues.
  • Perform RCA, impact assessments, and implement permanent corrective actions.
  • Design and enhance monitoring, observability, and runbooks to improve reliability.
  • Drive automation using scripting, IaC, CI/CD pipelines, and self-healing capabilities.
  • Partner with engineering, infrastructure, and product teams to remediate recurring issues.
  • Serve as Incident Commander during major incidents and coordinate cross-functional response.
  • Provide leadership and workload management for SRE, DevOps, and production support engineers.
  • Track metrics (MTTR, MTTD, SLA) to drive continuous improvement.

Skills

Leadership
Incident management
Stakeholder communication

Education

Bachelor's degree or equivalent

Tools

AWS
Azure
Kubernetes
Docker
GitHub Actions
Azure DevOps
Jenkins
GitLab
Datadog
Splunk
Dynatrace
Grafana
Prometheus
CloudWatch
Azure Monitor
OpenTelemetry
ServiceNow
Jira
Terraform
Ansible
REST APIs
SQL

Job description

U.S. Bank in Chicago seeks a senior Site Reliability Engineer to lead complex production incidents, perform RCA, and drive reliability across cloud platforms. You will own the incident command during major outages and collaborate with software, infra, and product teams to implement lasting fixes.

Hands-on with AWS/Azure, Kubernetes, Docker, CI/CD, and observability tools is required. You will mentor engineers, optimize monitoring and runbooks, and use MTTR/MTTD metrics to push continuous

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE Lead: Incident Commander & Reliability Champion
SRE Lead: Incident Commander & Reliability Champion

U.S. Bank • Northern (KY)

Hybrid
USD 112,000 - 131,000
Healthcare
Retirement plan
Paid vacation
+2
SRE Lead: Incident Commander for Cloud Reliability
SRE Lead: Incident Commander for Cloud Reliability

Us Bank • Atlanta (GA)

On-site
USD 112,000 - 131,000
Healthcare
Life Insurance
Disability
+6
Senior Incident Command & Reliability Engineer
Senior Incident Command & Reliability Engineer

IBM • Boston (MA)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer - AWS, Java & Kubernetes
Senior Site Reliability Engineer - AWS, Java & Kubernetes

JPMorganChase • Chicago (IL)

On-site
USD 130,000 - 190,000
Site Reliability Engineer: SRE, Monitoring & Observability
Site Reliability Engineer: SRE, Monitoring & Observability

U.S. Bank • Irving (TX)

On-site
USD 105,000 - 124,000
Healthcare (medical, dental, vision)
401(k) and employer-funded retirement
Senior Site Reliability Engineer - AWS, Java & Kubernetes
Senior Site Reliability Engineer - AWS, Java & Kubernetes

JPMorgan Chase & Co. • Chicago (IL)

On-site
USD 120,000 - 180,000
Senior SRE Lead: Cloud Platform Reliability & Automation
Senior SRE Lead: Cloud Platform Reliability & Automation

Bank of America • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Senior Site Reliability Engineer - Cloud & Automation Lead
Senior Site Reliability Engineer - Cloud & Automation Lead

Ss • Waltham (MA)

Hybrid
USD 130,000 - 140,000
Senior Site Reliability Engineer - Cloud & Incidents
Senior Site Reliability Engineer - Cloud & Incidents

Habitat For Humanity Of Durham • Raleigh (NC)

On-site
USD 120,000 - 155,000
Senior Cloud DevOps Engineer - AWS, Terraform & Networking
Senior Cloud DevOps Engineer - AWS, Terraform & Networking

Us Bank • Chicago (IL)

On-site
USD 124,000 - 146,000
Healthcare
401(k)
Paid Vacation
+6