Site Reliability Engineer

Latent Space

San Francisco (CA)

On-site

USD 160,000 - 210,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Competitive equity
Medical, dental, vision insurance
Flexible PTO
Paid parental leave
Carrot fertility stipend

Job summary

Latent is hiring a Site Reliability Engineer to own the production environment for our clinical AI platform. You will design and maintain scalable, high‑availability systems used by major health systems, leveraging Kubernetes, Helm, and Terraform.

You will optimize deployment pipelines (TypeScript and Python/ML), improve developer experience, and ensure reliable operations in a fast‑paced, in‑office SF culture. This role emphasizes ownership and rapid, safe shipping.

Qualifications

  • IaC & Orchestration: Kubernetes, Helm, Terraform expertise.
  • Scaling distributed systems with high availability experience.
  • Deployment experience for TypeScript apps and Python/ML models.
  • Core team member willing to work in SF office five days.

Responsibilities

  • Own production environment and ensure 99.9%+ stability.
  • Design, implement, and maintain infrastructure for 500+ deployments.
  • Manage containerized infra with Kubernetes and Helm.
  • Optimize CI/CD and deployment pipelines for reliability.
  • Maintain IaC with Terraform and improve DevX workflows.

Skills

Kubernetes
Helm
Terraform
CI/CD pipelines
DevX tooling
Automation
Linux CLI
Python/ML deployment
TypeScript deployment

Tools

Kubernetes
Helm
Terraform
PostgreSQL
Redis
Kafka
CI/CD pipelines
Python/ML deployment
TypeScript deployment

Job description

About Latent

Latent is the enterprise pharmacy intelligence platform. Our clinical AI streamlines prior authorizations, appeals, and 340B compliance, so patients start therapy faster and care teams spend less time on paperwork.

We raised an $80M Series A led by Spark Capital and Transformation Capital, with General Catalyst, McKesson Ventures, Conviction, and Y Combinator. 60+ health systems run on Latent, including Yale New Haven, Mount Sinai, UCSF, and Ochsner. We move fast, operate with high ownership, and build products that directly improve patient care.

The Role

You are the infrastructure expert who enables our rapid product development and guarantees 99.9%+ stability and performance of our clinical AI platform for major health systems. Your focus on operational excellence is directly tied to a patient's access to life‑saving treatment.

What You’ll Do

As our SRE, you will own the entire production environment and improve the development experience:

  • Infrastructure Ownership: Design, implement, and maintain the production environment, having previously handled 500+ machine deployments.
  • Kubernetes Mastery: Own our containerized infrastructure, leveraging deep expertise in Kubernetes and Helm to manage deployment, scaling, and operational health.
  • CI/CD & Deployment Optimization: Optimize and streamline both the TypeScript and Python/ML deployment pipelines to support high‑velocity feature release while maintaining the highest reliability.
  • DevX Support: Support Developer Experience (DevX) work to streamline developer workflows, enhance tool proficiency, and improve CI/CD systems.
  • Infrastructure as Code (IaC): Manage and maintain infrastructure definitions using Terraform.
What We’re Looking For

You have the intensity and technical mastery to own mission‑critical infrastructure. You hold yourself and others to high standards and thrive in a high‑energy, in‑office culture where everyone is in it to win it.

  • Tool Proficiency: You are highly proficient with your tools—you speak command line fluently and have mastered keyboard shortcuts.
  • Ownership: You thrive on owning complex systems and have a proven track record of scaling mission‑critical deployments.
  • Automation Drive: You love automating things, always finding new ways to increase your own leverage, and defining standards for operational excellence.
  • Problem Solver: You won’t wait for someone else to solve a problem that you’re in a position to solve; you are willing to jump into whatever needs to get done.
Technical Qualifications & Environment
  • IaC & Orchestration: Deep, demonstrable experience with Kubernetes, Helm, and Terraform.
  • Scaling Systems: Proven ability to architect and maintain complex, distributed systems with high‑availability requirements.
  • Deployment Experience: Hands‑on experience optimizing deployment pipelines for both application code (TypeScript) and machine learning models (Python/ML). Also PostgreSQL, Redis, Kakfa.
  • Core Team Member: Excitement about working five days per week in our San Francisco office.
Nice to Have
  • Interest in healthcare, pharmacy, or health technology
  • Uses AI tools in your workflow and is always looking for ways to move faster with them
Why Join Latent
  • Backed by top‑tier investors including General Catalyst, Conviction, and Y Combinator
  • Work alongside a high‑caliber team building products our healthcare partners love
  • Work on mission‑critical problems at the intersection of AI and healthcare
  • Real ownership and visibility
  • High‑impact role on a small, fast‑growing team
Benefits
  • Competitive compensation, including meaningful equity
  • Medical, dental, and vision insurance for employee and dependents
  • Flexible PTO policy
  • Paid parental leave
  • Fertility and family‑building stipend through Carrot
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer (Frontend)
Software Engineer (Frontend)

Latent Space • San Francisco (CA)

On-site
USD 150,000 - 210,000
Equity
Healthcare benefits
Flexible PTO
+2
Machine Learning Engineer
Machine Learning Engineer

Latent Space • San Francisco (CA)

On-site
USD 170,000 - 210,000
Equity
Medical insurance
Dental insurance
+4
Research Scientist
Research Scientist

Latent Space • San Francisco (CA)

On-site
USD 180,000 - 240,000
Equity
Health insurance
Flexible PTO
+2
Senior Product Manager
Senior Product Manager

Latent • San Francisco (CA)

On-site
USD 140,000 - 190,000
Competitive compensation
Medical, dental, and vision insurance
Flexible PTO policy
+2
Principal Technical Sourcer
Principal Technical Sourcer

Latent Space • San Francisco (CA)

On-site
USD 90,000 - 140,000
Equity
Medical, dental, and vision insurance
Flexible PTO
+2
Strategic Account Executive (New York City)
Strategic Account Executive (New York City)

LATENT • New York (NY)

On-site
USD 180,000 - 320,000
Competitive compensation
Meaningful equity
Medical, dental, and vision insurance
+3
Product Design San Francisco →
Product Design San Francisco →

Latent, Inc. • New York (NY)

On-site
USD 120,000 - 180,000
Competitive pay
Meaningful equity
Health insurance
+3
Account Manager - Health Systems (San Francisco)
Account Manager - Health Systems (San Francisco)

Latent • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive compensation
Meaningful equity
Medical, dental, and vision insurance
+3
Product Design
Product Design

Latent Space • San Francisco (CA)

On-site
USD 130,000 - 180,000
Competitive compensation with equity
Medical, dental, and vision insurance
Flexible PTO
+2
Product Manager
Product Manager

Latent • San Francisco (CA)

On-site
USD 110,000 - 150,000