Staff SRE: Release Engineering & Reliability

Plaid

San Francisco (CA)

On-site

USD 207,600 - 273,600

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Equity
Commission

Job summary

Plaid is seeking a Staff Site Reliability Engineer on Release Engineering to scale reliability across product engineering. You will architect the SLO and error-budget programs, drive progressive delivery, and ensure new products are production-ready.

By partnering with Platform and Infrastructure teams, you will translate complex production needs into intuitive, self-service tooling, lead incident responses, and shape deployment health for faster, safer software delivery.

Qualifications

  • Over 8 years of professional experience in backend systems, SRE, or platform engineering.
  • Proven track record designing reliability programs with cross-team adoption.
  • Experience building or operating canary rollout systems, metric-gated analysis, or automated rollback infrastructure.
  • Strong Go or similar systems language proficiency.

Responsibilities

  • Lead the expansion of reliability standards across product engineering and tooling.
  • Architect and manage the SLO and error-budget framework for strategic product choices.
  • Promote progressive delivery and automated safety gates to maintain velocity with stability.
  • Guide product teams toward production readiness with observability and incident response expertise.
  • Collaborate with Platform and Infrastructure teams to translate complex requirements into self-service tooling.
  • Direct responses to critical incidents and ensure post-mortem actions improve the platform.
  • Prepare for an AI-driven development landscape by scaling safety nets for more frequent code changes.

Skills

SRE / Platform engineering
Go language
Observability
Incident response
Platform engineering
Canary deployments

Tools

Kubernetes
Prometheus
ArgoCD
Service mesh

Job description

Plaid is seeking a Staff Site Reliability Engineer on Release Engineering to scale reliability across product engineering. You will architect the SLO and error-budget programs, drive progressive delivery, and ensure new products are production-ready.

By partnering with Platform and Infrastructure teams, you will translate complex production needs into intuitive, self-service tooling, lead incident responses, and shape deployment health for faster, safer software delivery.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer, Release Engineering
Staff Site Reliability Engineer, Release Engineering

Plaid Inc • New York (NY)

Hybrid
USD 130,000 - 180,000
Comprehensive benefit plan (medical, dental, vision, 401(k))
Equity and/or commission depending on position
Equal opportunity employer
Remote Staff SRE: Reliability, Incident Recovery & Growth
Remote Staff SRE: Reliability, Incident Recovery & Growth

Fingerprint • United States

Remote
USD 150,000 - 210,000
Staff SRE: Scale, Observability & Automation Leader
Staff SRE: Scale, Observability & Automation Leader

Replit • Northern (KY)

Hybrid
USD 180,000 - 260,000
Salary & equity
401(k) matching
Health, dental, vision, life
+9
Staff SRE: Platform Reliability Architect
Staff SRE: Platform Reliability Architect

Anduril Industries • Costa Mesa (CA)

On-site
USD 191,000 - 253,000
Equity grants
Comprehensive benefits
Senior SRE: Observability, Automation & Scalable Systems
Senior SRE: Observability, Automation & Scalable Systems

Replit • Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive Salary & Equity
401(k) 4% match (US)
Health, Dental, Vision & Life
+7
Senior SRE: Scale Reliability Leader (Hybrid)
Senior SRE: Scale Reliability Leader (Hybrid)

Plenful • San Francisco (CA)

Hybrid
USD 180,000 - 250,000
Healthcare Coverage
401(k) with Company Match
Equity
+5
Senior SRE: Production Reliability & Observability
Senior SRE: Production Reliability & Observability

Stradit LLC • Dallas (TX), Northern (KY)

Hybrid
USD 140,000 - 190,000
Staff SRE: Scale, Observability & Kubernetes
Staff SRE: Scale, Observability & Kubernetes

Replit • Foster City (CA)

On-site
USD 180,000 - 260,000
Competitive Salary & Equity
401(k) with 4% match
Health, Dental, Vision and Life Ins.
+2
SRE Engineer — Scale AI Platforms (Hybrid Work)
SRE Engineer — Scale AI Platforms (Hybrid Work)

Plaud • San Francisco (CA)

Hybrid
USD 180,000 - 230,000
ESOP
Hybrid work model
Health benefits
Global SRE Leader: Platform Reliability & Automation
Global SRE Leader: Platform Reliability & Automation

Broadridge Financial Solutions • New York (NY)

Hybrid
USD 235,000 - 250,000