Senior Site Reliability Engineer - Healthcare Infra Equity

Enzo Health

Lehi (UT)

On-site

USD 120,000 - 180,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive salary
Equity
401k & Insurance
High ownership
Direct collaboration with founders

Job summary

Enzo Health is hiring a Senior Site Reliability Engineer to join our Security and Site Reliability team. You will focus on a stable, scalable AWS and Kubernetes platform, reliable Postgres operations, and safer delivery. You will mature observability, on‑call, and incident response practices across engineering.

This is a hands‑on role. You will diagnose production problems, improve infrastructure, write automation, and help product engineers operate their services with confidence.

Qualifications

  • At least 5 years of experience in site reliability, platform, infrastructure, or production engineering.
  • Strong hands‑on experience with AWS and production Kubernetes.
  • Strong experience with Terraform and infrastructure as code.
  • Experience operating Postgres in production, including backup and restore, performance, and safe migrations.
  • Experience building and operating CI/CD and deployment systems.
  • Experience with modern observability tools across metrics, logs, and traces.
  • Strong programming or scripting skills for automation and operational tools.
  • A proven track record diagnosing production failures and leading incidents through recovery.
  • Ability to write clear automation, runbooks, and technical standards.
  • Ability to work independently and partner well with application engineers.

Responsibilities

  • Operate and improve our AWS and Kubernetes environments.
  • Own infrastructure changes through Terraform, including modules, state, review standards, and drift control.
  • Improve environment management, cluster practices, and deployment reliability.
  • Make CI/CD and releases safer through clear checks, reliable rollbacks, environment consistency, and release observability.
  • Improve capacity planning, resource controls, resilience, and cost visibility.
  • Build automation that removes repetitive operational work and reduces avoidable failures.
  • Own the operational health of Postgres in production.
  • Verify backups and regularly test restore procedures against agreed recovery targets.
  • Improve monitoring for connections, storage, slow queries, locks, and other important failure signals.
  • Guide safe schema migrations, database access, and production change procedures.
  • Find performance and capacity risks before they affect customers.
  • Maintain clear database runbooks for common failures and emergency work.
  • Mature dashboards, logs, traces, and alert routing.
  • Validate service‑level indicators and objectives; close coverage gaps.
  • Participate in on‑call rotation and write runbooks for engineers.
  • Lead or support incident response and blameless post‑incident reviews.
  • Use reliability data to set priorities and measure improvement.

Skills

AWS
Kubernetes
Terraform
Postgres
CI/CD
Observability
Automation
Scripting
Incident response

Tools

GitHub Actions

Job description

Enzo Health is hiring a Senior Site Reliability Engineer to join our Security and Site Reliability team. You will focus on a stable, scalable AWS and Kubernetes platform, reliable Postgres operations, and safer delivery. You will mature observability, on‑call, and incident response practices across engineering.

This is a hands‑on role. You will diagnose production problems, improve infrastructure, write automation, and help product engineers operate their services with confidence.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE - HealthTech Reliability & Security
Senior SRE - HealthTech Reliability & Security

Tandem Inc. • Lehi (UT), Northern (KY)

Hybrid
USD 120,000 - 180,000
Competitive salary
Meaningful equity
401k & insurance
+1
Site Reliability Engineer
Site Reliability Engineer

Tandem Inc. • Lehi (UT), Northern (KY)

On-site
USD 120,000 - 180,000
Competitive salary
Meaningful equity
401k & insurance
+1
Site Reliability Engineer
Site Reliability Engineer

Enzo Health • Lehi (UT)

On-site
USD 120,000 - 180,000
Competitive salary
Equity
401k & Insurance
+2
Senior Staff Site Reliability Engineer
Senior Staff Site Reliability Engineer

Pivotal Health • Santa Monica (CA)

Hybrid
USD 180,000 - 250,000
Competitive compensation
Health, dental, and vision coverage
401(k) retirement plan
+2
Senior Site Reliability Engineer: Scalable Health Platform
Senior Site Reliability Engineer: Scalable Health Platform

Socket.dev • Williamsburg (VA)

Hybrid
USD 180,000 - 240,000
Competitive compensation including equ
Health coverage
401(k) retirement plan
+2
Staff Site Reliability Engineer - Scale Global AWS Infra
Staff Site Reliability Engineer - Scale Global AWS Infra

Pearl Street Technologies • Pittsburgh

On-site
USD 130,000 - 180,000
Senior DevOps Engineer for Healthcare AI Platform SRE
Senior DevOps Engineer for Healthcare AI Platform SRE

Transformcap • Palo Alto (CA)

Hybrid
USD 170,000 - 220,000
Equity
Medical insurance
Flexible hours
+1
Site Reliability Engineer — Healthcare Platform & Security
Site Reliability Engineer — Healthcare Platform & Security

Artha Nexgen • Jersey City (NJ), Northern (KY)

On-site
USD 90,000 - 130,000
Free lunch
Unlimited PTO
Health benefits for employees
+3
Senior DevOps/SRE for Healthcare AI Platform
Senior DevOps/SRE for Healthcare AI Platform

eSolutionsFirst • Palo Alto (CA), Northern (KY)

Hybrid
USD 170,000 - 220,000
Equity
Senior Site Reliability Engineer - Cloud & Kubernetes
Senior Site Reliability Engineer - Cloud & Kubernetes

Fabric • New York (NY)

On-site
USD 140,000 - 190,000