Lead Site Reliability Engineer, Chief Digital Office

Worky

Eden Prairie (MN)

Hybrid

USD 113,000 - 193,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Remote work

Job summary

UnitedHealth Group's OptumRx Digital is seeking a Lead Site Reliability Engineer to own the reliability, performance, and resilience of critical pharmacy services and cloud infrastructure. You will drive measurable improvements in availability and customer experience across high‑throughput digital platforms.

You will mentor engineers, define SRE best practices, and lead incident response, RCA, and post‑mortems, while partnering with product and software teams to optimize the digital supply

Qualifications

  • 7+ years of professional experience in Site Reliability Engineering, DevOps, or Software Engineering supporting enterprise applications and cloud infrastructure.
  • 5+ years of experience managing, monitoring, and maintaining production cloud applications and microservices (e.g., AWS, Azure, or GCP)
  • 4+ years of experience with containerized application environments and orchestration frameworks (e.g., Docker, Kubernetes)
  • 3+ years of experience setting up application performance monitoring (APM) and enterprise observability platforms (e.g., Datadog, Dynatrace, Prometheus, Grafana, or Splunk)

Responsibilities

  • Drive overall technical accountability for the availability, performance, and operational well-being of OptumRx Digital applications, microservices, and supporting infrastructure
  • Partner with software engineering and product leadership to influence digital supply chain technology workflows, optimizing transactional throughput, system integrations, and application health
  • Architect and implement robust application performance monitoring (APM), logging, and observability solutions to proactively detect, diagnose, and resolve application and service degradations
  • Automate cloud infrastructure, deployment pipelines, and operational processes using Infrastructure as Code (IaC) and modern CI/CD practices
  • Establish, track, and champion key application reliability metrics, including Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets across applications and services
  • Lead end-to-end incident management, root cause analysis (RCA), and post-mortem actions to drive continuous improvement and eliminate recurring application failures across supply chain platforms
  • Provide technical leadership, mentorship, and guidance to engineering teams, fostering a culture of operational rigor, engineering quality, and continuous delivery

Skills

Site Reliability Engineering
DevOps
Cloud infrastructure
Observability
Leadership

Tools

Docker
Kubernetes
Datadog
Dynatrace
Prometheus
Grafana
Splunk
Terraform
CloudFormation
Ansible

Job description

Optum Tech is a global leader in health care innovation. Our teams develop cutting-edge solutions that help people live healthier lives and help make the health system work better for everyone. From advanced data analytics and AI to cybersecurity, we use innovative approaches to solve some of health care’s most complex challenges. Your contributions here have the potential to change lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.

Position Summary

As a Lead Site Reliability Engineer within OptumRx Digital, you will be accountable for the health, performance, operational resilience, and overall well-being of our critical digital applications, pharmacy services, and supporting cloud infrastructure. In this role, you will drive operational excellence across high-throughput digital platforms, ensuring maximum reliability and seamless customer experiences for millions of patients and pharmacy partners. You will play a pivotal role in influencing and optimizing our digital supply chain technology workflows, partnering closely with engineering, product, and operational leadership to build resilient software systems, streamline deployment pipelines, and proactively safeguard application health.

You’ll enjoy the flexibility to work remotely * from anywhere within the U.S. as you take on some tough challenges.

For all hires in the Minneapolis or Washington, D.C. area, you will be required to work in the office a minimum of four days per week.

Primary Responsibilities:
  • Drive overall technical accountability for the availability, performance, and operational well-being of OptumRx Digital applications, microservices, and supporting infrastructure
  • Partner with software engineering and product leadership to influence digital supply chain technology workflows, optimizing transactional throughput, system integrations, and application health
  • Architect and implement robust application performance monitoring (APM), logging, and observability solutions to proactively detect, diagnose, and resolve application and service degradations
  • Automate cloud infrastructure, deployment pipelines, and operational processes using Infrastructure as Code (IaC) and modern CI/CD practices
  • Establish, track, and champion key application reliability metrics, including Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets across applications and services
  • Lead end-to-end incident management, root cause analysis (RCA), and post-mortem actions to drive continuous improvement and eliminate recurring application failures across supply chain platforms
  • Provide technical leadership, mentorship, and guidance to engineering teams, fostering a culture of operational rigor, engineering quality, and continuous delivery

You’ll be rewarded and recognized for your performance in an environment that will challenge you and give you clear direction on what it takes to succeed in your role as well as provide development for other roles you may be interested in.

Required Qualifications:
  • 7+ years of professional experience in Site Reliability Engineering, DevOps, or Software Engineering supporting enterprise applications and cloud infrastructure
  • 5+ years of experience managing, monitoring, and maintaining production cloud applications and microservices (e.g., AWS, Azure, or GCP)
  • 4+ years of experience with containerized application environments and orchestration frameworks (e.g., Docker, Kubernetes)
  • 3+ years of experience setting up application performance monitoring (APM) and enterprise observability platforms (e.g., Datadog, Dynatrace, Prometheus, Grafana, or Splunk)
Preferred Qualifications:
  • Relevant technical certifications such as AWS/Azure/GCP Solutions Architect or Certified Kubernetes Administrator (CKA)
  • 4+ years of experience implementing Infrastructure as Code (IaC) using tools such as Terraform, CloudFormation, or Ansible
  • 3+ years of experience writing scripts or code in languages such as Python, Go, Java, or Nodejs to automate application and infrastructure operations
  • Experience supporting digital pharmacy, healthcare supply chain, e-commerce, or high-volume transactional application ecosystems
  • Proven track record of influencing cross-functional technical architectures to improve end-to-end application health and supply chain resiliency
  • Deep understanding of high-volume event streaming, API gateway management, and distributed database performance optimization

*All employees working remotely will be required to adhere to UnitedHealth Group’s Telecommuter Policy

Pay is based on several factors including but not limited to local labor markets, education, work experience, certifications, etc. In addition to your salary, we offer benefits such as, a comprehensive benefits package, incentive and recognition programs, equity stock purchase and 401k contribution (all benefits are subject to eligibility requirements). No matter where or when you begin a career with us, you’ll find a far-reaching choice of benefits and incentives. The salary for this role will range from $112,700 - $193,200 annually based on full-time employment. We comply with all minimum wage laws as applicable.

Application Deadline:

This will be posted for a minimum of 2 business days or until a sufficient candidate pool has been collected. Job posting may come down early due to volume of applicants.

At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.

UnitedHealth Group is an Equal Employment Opportunity employer under applicable law and qualified applicants will receive consideration for employment without regard to race, national origin, religion, age, color, sex, sexual orientation, gender identity, disability, or protected veteran status, or any other characteristic protected by local, state, or federal laws, rules, or regulations.

UnitedHealth Group is a drug - free workplace. Candidates are required to pass a drug test before beginning employment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AWS Cloud Site Reliability Engineer
AWS Cloud Site Reliability Engineer

JobCubby • Northern (KY)

Hybrid
USD 73,000 - 130,000
Principal Site Reliability Engineer - Remote
Principal Site Reliability Engineer - Remote

UnitedHealth Group • Eden Prairie (MN)

Hybrid
Confidential
Remote work options
Lead Software Engineer - Remote
Lead Software Engineer - Remote

Optum • Schaumburg (IL)

Remote
USD 112,700 - 193,200
Senior Software Engineer
Senior Software Engineer

UnitedHealth Group • Minnetonka (MN), Northern (KY)

Hybrid
Confidential
PTO & Holidays
Medical & Dental Coverage
401(k) plan
+5
SVP, Digital Identity and Tech Products
SVP, Digital Identity and Tech Products

Optum • Eden Prairie (MN)

Hybrid
USD 225,000 - 375,000
Technical Program Lead
Technical Program Lead

UnitedHealth Group • Mesa (AZ)

Hybrid
Confidential
Comprehensive benefits package
Equity stock purchase
401k contribution
Senior Digital Product Management Consultant - Remote
Senior Digital Product Management Consultant - Remote

NHSHP • Eden Prairie (MN), Northern (KY)

Hybrid
USD 92,000 - 164,000
Sr Mgr Software Engineering
Sr Mgr Software Engineering

Optum • Eden Prairie (MN)

On-site
USD 113,000 - 193,000
Comprehensive benefits
Equity stock purchase
401k contribution
Senior Digital Product Management Consultant - Remote
Senior Digital Product Management Consultant - Remote

Texas Health Institute • Eden Prairie (MN), Northern (KY)

Hybrid
USD 92,000 - 164,000
Director, Payer Technology Optum Advisory - Remote
Director, Payer Technology Optum Advisory - Remote

Texas Health Institute • Eden Prairie (MN), Northern (KY)

Hybrid
USD 135,000 - 231,000