Senior SRE: Cloud-Native Reliability & Observability

UnitedHealth Group

Schaumburg (IL)

Remote

USD 92,000 - 164,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Telecommute within US

Job summary

UnitedHealth Group is seeking an experienced Site Reliability Engineer to lead and evolve our SRE practices across cloud-native platforms. You will design and implement scalable, highly available infrastructure and monitoring solutions, while guiding incident response and reliability improvements.

Your role involves mentoring engineers, building IaC, and collaborating with cross-functional teams to advance platform resilience and automation across multi-cloud environments.

Qualifications

  • 7+ years in Site Reliability Engineering, DevOps, Platform or Software Engineering

Responsibilities

  • Lead implementation and improvement of SRE practices including SLIs, SLOs and error budgets
  • Design, deploy and support cloud-native platforms across AWS, Azure, and Kubernetes
  • Build observability solutions with Datadog, Splunk, Grafana and OpenTelemetry
  • Develop dashboards, alerting, health scorecards and reliability metrics
  • Lead incident response, RCA and post-incident remediation
  • Create IaC using Terraform and cloud automation
  • Enable automation and self-healing to reduce toil
  • Collaborate with software engineering, security and platform teams
  • Support Kubernetes-based platforms and microservices
  • Implement CI/CD and GitOps with GitHub Actions, ArgoCD, Azure DevOps
  • Drive AI-enabled operational capabilities and anomaly detection
  • Mentor engineers on SRE and cloud-native best practices

Skills

SRE principles
Observability
Cloud engineering
CI/CD
Python scripting
Automation
On-call experience
Mentoring

Education

Bachelor's degree in CS/Eng/IT

Tools

Kubernetes
Datadog
Splunk
Grafana
OpenTelemetry
Prometheus
Terraform
GitHub Actions
ArgoCD
Azure DevOps

Job description

UnitedHealth Group is seeking an experienced Site Reliability Engineer to lead and evolve our SRE practices across cloud-native platforms. You will design and implement scalable, highly available infrastructure and monitoring solutions, while guiding incident response and reliability improvements.

Your role involves mentoring engineers, building IaC, and collaborating with cross-functional teams to advance platform resilience and automation across multi-cloud environments.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE: Cloud-Native Reliability & Observability
Senior SRE: Cloud-Native Reliability & Observability

Optum • Schaumburg (IL)

Remote
USD 92,000 - 164,000
Senior SRE: Infrastructure Observability & Reliability
Senior SRE: Infrastructure Observability & Reliability

NewGen Technologies • Owings Mills (MD), Northern (KY)

Hybrid
USD 150,000 - 190,000
Remote Cloud SRE: Build Resilient Health-Tech Infra
Remote Cloud SRE: Build Resilient Health-Tech Infra

UnitedHealth Group • Eden Prairie (MN)

Remote
Confidential
Remote work flexibility
Comprehensive benefits package
Equity stock purchase
+1
Senior Platform DevOps Engineer – Remote SRE
Senior Platform DevOps Engineer – Remote SRE

UnitedHealth Group • Eden Prairie (MN)

Hybrid
USD 92,000 - 164,000
Comprehensive benefits package
Equity stock purchase plan
401(k) contribution
Senior SRE - Observability & Cloud Reliability
Senior SRE - Observability & Cloud Reliability

Cisco Systems, Inc. • San Francisco (CA)

On-site
USD 168,000 - 245,000
Medical benefits
401(k) matching
Parental leave
+1
Senior SRE - Hybrid, Observability & Reliability
Senior SRE - Hybrid, Observability & Reliability

Early Warning Services LLC • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Plan with match
PTO and Holidays
+1
Remote Lead SRE - Cloud & Reliability
Remote Lead SRE - Cloud & Reliability

UnitedHealth Group • Eden Prairie (MN)

Hybrid
USD 113,000 - 193,000
Remote work
Senior SRE Lead: Reliability, Observability & Cloud
Senior SRE Lead: Reliability, Observability & Cloud

Shieldai • San Mateo (CA)

On-site
USD 180,000 - 240,000
Senior SRE: Automate Reliability & Observability
Senior SRE: Automate Reliability & Observability

Bank of America • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Senior SRE: Observability & Automation Leader (Hybrid)
Senior SRE: Observability & Automation Leader (Hybrid)

ISO New England • Holyoke (MA)

Hybrid
USD 134,000 - 170,000
Hybrid work environment
Relocation assistance
Tuition reimbursement
+6