AI Platform DevOps & SRE Lead

Reactor

San Francisco (CA)

On-site

USD 100,000 - 160,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive salary and early equity
Visa sponsorship
Generous health, dental, and vision coverage

Job summary

Reactor is looking for a DevOps/SRE engineer in San Francisco to enhance the reliability and observability of their AI platform. The position requires running production Kubernetes clusters, strong CI/CD pipeline experience, and a solid understanding of GitOps and infrastructure as code. Successful candidates will triage production issues, manage secret infrastructure, and define SLOs. Benefits include a competitive salary, equity, and health coverage, with opportunities for relocation support.

Qualifications

  • Proven experience in running production Kubernetes clusters.
  • Strong background in CI/CD practices and tools.
  • Expert in GitOps and reconciliations.
  • Fluency in Infrastructure as Code with Terraform.
  • Experience in setting up observability and defining SLOs.
  • Proficiency in secret management practices.
  • Experience in incident response with effective postmortems.
  • Ability to write scripts in Go, Python, or Bash.

Responsibilities

  • Own and evolve CI/CD pipelines and deployment lifecycles.
  • Maintain observability stack for all services.
  • Define SLOs and manage incident response.
  • Handle infrastructure as code across different cloud providers.
  • Operate secret management and ensure deployment safety.

Skills

Kubernetes management
CI/CD pipeline development
GitOps practices
Infrastructure as Code
Observability and monitoring
Incident response
Coding proficiency in Go, Python, or Bash
Secret management

Tools

Terraform
Helm

Job description

Reactor is looking for a DevOps/SRE engineer in San Francisco to enhance the reliability and observability of their AI platform. The position requires running production Kubernetes clusters, strong CI/CD pipeline experience, and a solid understanding of GitOps and infrastructure as code. Successful candidates will triage production issues, manage secret infrastructure, and define SLOs. Benefits include a competitive salary, equity, and health coverage, with opportunities for relocation support.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff, DevOps
Member of Technical Staff, DevOps

Reactor • San Francisco (CA)

On-site
USD 100,000 - 160,000
Competitive salary and early equity
Visa sponsorship
Generous health, dental, and vision coverage
Senior DevOps Engineer—AI Platform Reliability
Senior DevOps Engineer—AI Platform Reliability

Flux Enterprise • San Francisco (CA)

On-site
USD 120,000 - 150,000
Senior SRE: AI-Driven Kubernetes Reliability at Scale
Senior SRE: AI-Driven Kubernetes Reliability at Scale

fal - Features & Labels • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health insurance
Dental insurance
Vision insurance
+1
Senior Platform Engineer (SRE) - AI Control Plane
Senior Platform Engineer (SRE) - AI Control Plane

Speakeasy • San Francisco (CA)

On-site
USD 120,000 - 160,000
Staff Backend Engineer: AI-Driven SRE & Scalable Systems
Staff Backend Engineer: AI-Driven SRE & Scalable Systems

Resolve AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Comprehensive Medical, Dental, and Vision Insurance
Monthly Housing Stipend
Flexible (Unlimited) Paid Time Off
+5
AI Platform SRE: Reliability, Observability & Scale
AI Platform SRE: Reliability, Observability & Scale

Schonfeld • New York (NY)

On-site
USD 175,000 - 225,000
Platform SRE Lead – AI Spacetech, Equity & Ownership
Platform SRE Lead – AI Spacetech, Equity & Ownership

Attis • Houston (TX)

On-site
USD 130,000 - 150,000
Flexible time-off policy
Cost-effective healthcare
401k matching plan
+1
SRE: Scalable ML Infra & CI/CD Architect
SRE: Scalable ML Infra & CI/CD Architect

Baseten • San Francisco (CA)

On-site
USD 165,000 - 330,000
Competitive compensation with equity
100% medical, dental, and vision coverage
Generous PTO including Winter Break
+2
Senior AI Platform SRE: Scale Cloud Infra & Kubernetes
Senior AI Platform SRE: Scale Cloud Infra & Kubernetes

GCS Recruitment • Mount Laurel Township (NJ)

On-site
USD 110,000 - 170,000
SRE & Platform Engineer — Automate Reliability & Releases
SRE & Platform Engineer — Automate Reliability & Releases

Getpeer • San Francisco (CA), Northern (KY)

Hybrid
USD 190,000 - 240,000
Meaningful equity grant
Remote-first with quarterly San Francs
Health coverage
+3