Senior SRE — AI-Native Cloud & Infra Leader

Formation Bio

San Francisco (CA)

Hybrid

USD 186,000 - 232,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Equity
Comprehensive benefits
Generous perks

Job summary

Formation Bio is seeking a Senior Site Reliability Engineer to build and operate the infrastructure, delivery systems, and operational practices that enable our engineering teams to ship reliable software quickly. You will work across cloud infrastructure, observability, ML workloads, and production systems in an AI-native environment.

You will partner with Product, Data Engineering, and Data Science to design scalable infrastructure, apply AI-assisted tooling for automation, and maintain robust

Qualifications

  • 5+ years of Site Reliability Engineering or similar discipline.
  • Production cloud infra experience with strong reliability judgment.
  • Diagnostics, incident response, root cause analysis, observability, and automation.
  • Experience with AWS and Snowflake; Azure/GCP/Vercel is a plus.
  • Docker, GitHub, Kubernetes, Python, Terraform/OpenTofu; Terragrunt a plus.
  • Experience supporting ML/AI workloads and MLOps infrastructure is a plus.
  • Strong collaboration across engineering, data science, and security.

Responsibilities

  • Own shared infra and platform for compute, runtimes, orchestration, and observability.
  • Build and operate secure, reliable infrastructure for product apps and ML workloads.
  • Develop core AWS infra and multi-cloud outposts across environments.
  • Create and optimize infrastructure as code, CI/CD pipelines, and reusable patterns.
  • Establish SLOs, monitoring, incident response, and post-incident follow-through.
  • Evaluate architecture with product, data engineering, and data science teams.
  • Use AI tools to accelerate development and automation while validating output.
  • Participate in on-call rotations and balance automation with ClickOps as needed.
  • Draft requirements, design docs, and mentor engineers on SRE fundamentals.

Skills

5+ years SRE
AWS
GCP/Azure
Docker
Kubernetes
Python
Terraform/OpenTofu
GitHub
Observability
CI/CD

Tools

Docker
GitHub
Kubernetes
Python
Terraform/OpenTofu
Terragrunt
AWS
Snowflake

Job description

Formation Bio is seeking a Senior Site Reliability Engineer to build and operate the infrastructure, delivery systems, and operational practices that enable our engineering teams to ship reliable software quickly. You will work across cloud infrastructure, observability, ML workloads, and production systems in an AI-native environment.

You will partner with Product, Data Engineering, and Data Science to design scalable infrastructure, apply AI-assisted tooling for automation, and maintain robust

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI-Native SRE Engineer
Senior AI-Native SRE Engineer

Formation • Town of Boston (NY)

Hybrid
USD 186,000 - 232,000
Equity
Benefits
AI-Native SRE: Cloud Infra for ML & Product Systems
AI-Native SRE: Cloud Infra for ML & Product Systems

Formation Bio • San Francisco (CA)

Hybrid
USD 186,000 - 232,000
Engineering Manager, AI Infrastructure & SRE
Engineering Manager, AI Infrastructure & SRE

Formation Bio • New York (NY)

Hybrid
USD 186,000 - 232,000
Engineering Manager: AI-Driven Infrastructure & SRE
Engineering Manager: AI-Driven Infrastructure & SRE

Formation Bio • San Francisco (CA)

Hybrid
USD 186,000 - 232,000
Equity
Comprehensive benefits
Hybrid work model (3 days in office)
Senior SRE - Cloud-Native ML Infra & AI Pipelines
Senior SRE - Cloud-Native ML Infra & AI Pipelines

Adobe • New York (NY)

On-site
USD 178,000 - 258,000
Senior SRE – AI Infrastructure Reliability Leader
Senior SRE – AI Infrastructure Reliability Leader

Nscale • San Francisco (CA), Seattle (WA), Houston (TX)

On-site
USD 170,000 - 265,000
Equity
Ownership from start
Flexible schedule
AI-Driven Infra Engineering Manager — SRE Lead (Hybrid)
AI-Driven Infra Engineering Manager — SRE Lead (Hybrid)

Scorpion Therapeutics • San Francisco (CA)

Hybrid
USD 170,000 - 210,000
Hybrid work model
AI-Native Infra & SRE Lead (Hybrid)
AI-Native Infra & SRE Lead (Hybrid)

Formation • Town of Boston (NY)

Hybrid
USD 186,000 - 232,000
Senior SRE: AI-Driven, Cloud-Native Reliability (Hybrid)
Senior SRE: AI-Driven, Cloud-Native Reliability (Hybrid)

OutSystems • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Hybrid work model
Senior SRE: AI-Driven Reliability & Automation (Hybrid)
Senior SRE: AI-Driven Reliability & Automation (Hybrid)

Namely • United States

Hybrid
USD 120,000 - 150,000