Staff Site Reliability Engineer

Manifold

Cambridge (MA)

On-site

USD 160,000 - 225,000

Full time

10 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Manifold is seeking a Staff Site Reliability Engineer to design, build, and operate the multi-account AWS infrastructure underpinning its platform. Youll work with Platform engineering and Professional Services to ensure secure, scalable, observable systems for internal and customer-facing use.

You will design infrastructure as code, automate deployments, manage CI/CD pipelines, and troubleshoot complex production issues while balancing automation with pragmatic approaches.

Qualifications

  • 7+ years in infrastructure, DevOps, SRE, or platform engineering with growing scope.
  • Hands-on cloud experience (AWS, GCP, or Azure) and troubleshooting across apps and infra.
  • Strong Terraform and CI/CD experience with modern tooling.
  • Familiar with identity systems, containers, and network tooling; data platform basics.

Responsibilities

  • Design and maintain infrastructure as code solutions across multi-account AWS.
  • Automate customer deployments, multi-account provisioning, and bootstrapping.
  • Manage CI/CD pipelines, build reliability, tests, and deployments.
  • Troubleshoot production issues across infra, data, and apps; leverage AI tools to reduce toil.
  • Address security/compliance in regulated environments.

Skills

SRE leadership
Cloud platforms
Terraform / IaC
CI/CD (Github Actions)
Security & compliance
Multi-account AWS environments
AI tooling awareness

Tools

Terraform
Github Actions
Docker
ECS
Okta
Auth0
Snowflake
Airflow
dbt
PostgreSQL
Tailscale
WireGuard
Packer

Job description

Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision medicine in the clinic, while maintaining the governance life sciences requires. Global companies and premier research institutions use Manifold to operate faster and more effectively. Backed by leading investors including Reach Capital, TQ Ventures, Calibrate Ventures, SilverArc Capital, and Industry Ventures, Manifold serves tens of thousands of users across hundreds of organizations globally.

About the Role

Manifold is looking for a Staff Site Reliability Engineer (SRE) to work at the intersection of AI, data infrastructure, and life sciences. In this high-impact role, you will help design, build, and operate the multi-account AWS infrastructure that acts as the foundation for Manifold’s platform.

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold’s internal and customer-facing infrastructure is secure, scalable, and observable. SREs are expected to be fluent across a wide-ranging tech stack, and comfortable working in a high-pressure, multi-threaded environment. They are also expected to balance a bias towards automation with more pragmatic approaches and have an intuitive understanding of the tradeoffs involved. The infrastructure you deploy and manage will play a key role in Manifold's success and help our customers bring life-changing medicines to patients faster.

What You'll Do
  • Design and maintain infrastructure as code solutions, thinking holistically about topology and component dependencies. You will have full responsibility for everything from Terraform plans to production observability.
  • Automate customer infrastructure deployments, including multi-account provisioning, database setup, workflow orchestration, and application bootstrapping.
  • Manage CI/CD pipelines, including build reliability, test stability, and deployment automation for Manifold services.
  • Troubleshoot complex production issues across infrastructure, data, and application layers. Leverage LLM to the fullest extend to minimize toil and and manage date to date operations.
  • Networking and security / compliance work required for highly regulated environments.
What You’ll Bring
  • 7+ years in infrastructure, DevOps, SRE, or platform engineering roles with increasing scope and autonomy. You are a leader / doer who can establish operational standards and drive technical direction while also staying hands-on.
  • Deep, hands-on cloud (AWS, GCP, or Azure) experience, hands-on application development experience, and comfortable in troubleshooting application issues.
  • Significant infrastructure-as-code (Terraform) experience. Strong CI/CD (Github Action) experience.
  • Familiarity with identity systems (Okta, Auth0), containerized deployments (Docker, ECS, Packer) and networking tooling (Tailscale, WireGuard). Working knowledge of data platform services, such as Snowflake, Airflow, dbt, and PostgreSQL.
  • Comfort managing complex, multi-account environments where customer isolation, security boundaries, and regulatory requirements add real constraints.
  • You move fast, make sound calls with the information available, and reset quickly when things don't go as planned. And you have strong bias towards pragmatic, incremental process automation.
  • You possess a track record of improving developer experience and reducing CI/CD friction.
  • The ability to effectively and positively collaborate with platform engineer, professional services, and customer IT groups.
  • You've built AI into how you work. You're curious about new tools, resourceful in applying them, and have concrete examples of how AI changed your output.
  • You've done your homework on what it means to accelerate life sciences research and you can articulate why that mission matters to you.
Why Manifold
  • We offer the unique opportunity to work at the frontier of AI in life sciences, helping the world's leading research and clinical institutions adopt transformative technology.
  • This is a high-autonomy, high-impact role where your work directly leads to positive customer outcomes that accelerate breakthrough science.
  • You will gain exposure to a wide range of technical and business challenges across oncology, genomics, clinical data, and AI agent development.
  • You will work with a collaborative team of engineers, scientists, and operators who are building something genuinely new and important.
  • This is an opportunity to grow fast in a company that takes talent seriously. We offer strong compensation packages and excellent benefits.
Salary Range

The base salary range for this position is $160,000–$225,000 annually.

  • Please note: Final offer amounts are determined by multiple factors, including prior experience and expertise, and may vary from the amount above. This range does not represent additional compensation benefits (such as equity, 401K match or medical, dental or vision insurance)
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Software Engineer | Core Platform
Staff Software Engineer | Core Platform

Manifold AI • Cambridge (MA)

On-site
USD 170,000 - 220,000
Manager, Strategic Accounts
Manager, Strategic Accounts

Manifold Ai • Cambridge (MA)

On-site
USD 140,000 - 180,000
Director Strategic Accounts | Broad Institute
Director Strategic Accounts | Broad Institute

Manifold Ai • Cambridge (MA)

On-site
USD 160,000 - 200,000
Equity
Competitive compensation
Engineering Manager | Terra Platform
Engineering Manager | Terra Platform

Manifold • Cambridge (MA)

Hybrid
USD 210,000 - 245,000
Equity
Competitive compensation
AI/ML Research Engineer
AI/ML Research Engineer

Manifold Bio • San Francisco (CA), Boston (MA)

On-site
USD 120,000 - 160,000
Senior/Staff Site Reliability Engineer - Data Center
Senior/Staff Site Reliability Engineer - Data Center

PathAI • Boston (MA)

On-site
USD 165,750 - 224,450
Staff Software Engineer (Full Stack)
Staff Software Engineer (Full Stack)

Prudentia Sciences • San Francisco (CA)

Hybrid
USD 230,000 - 275,000
Equity
Competitive salary
Remote-friendly
Site Reliability Engineer
Site Reliability Engineer

Instrumental Inc. • Palo Alto (CA)

On-site
USD 140,000 - 165,000
Health insurance
Vision insurance
Dental plan
+2
Staff Platform Engineer: Core Compute & Workflows
Staff Platform Engineer: Core Compute & Workflows

Manifold AI • Cambridge (MA)

On-site
USD 170,000 - 220,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Instrumental Inc. • Palo Alto (CA)

On-site
USD 175,000 - 229,000
Health benefits
Commuter plans
Parental leave