Staff SRE: Release Engineering & Reliability

Plaid

San Francisco (CA)

On-site

USD 207,600 - 273,600

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Commission

Job summary

Plaid is seeking a Staff Site Reliability Engineer on Release Engineering to scale reliability across product engineering. You will architect the SLO and error-budget programs, drive progressive delivery, and ensure new products are production-ready.

By partnering with Platform and Infrastructure teams, you will translate complex production needs into intuitive, self-service tooling, lead incident responses, and shape deployment health for faster, safer software delivery.

Qualifications

  • Over 8 years of professional experience in backend systems, SRE, or platform engineering.
  • Proven track record designing reliability programs with cross-team adoption.
  • Experience building or operating canary rollout systems, metric-gated analysis, or automated rollback infrastructure.
  • Strong Go or similar systems language proficiency.

Responsibilities

  • Lead the expansion of reliability standards across product engineering and tooling.
  • Architect and manage the SLO and error-budget framework for strategic product choices.
  • Promote progressive delivery and automated safety gates to maintain velocity with stability.
  • Guide product teams toward production readiness with observability and incident response expertise.
  • Collaborate with Platform and Infrastructure teams to translate complex requirements into self-service tooling.
  • Direct responses to critical incidents and ensure post-mortem actions improve the platform.
  • Prepare for an AI-driven development landscape by scaling safety nets for more frequent code changes.

Skills

SRE / Platform engineering
Go language
Observability
Incident response
Platform engineering
Canary deployments

Tools

Kubernetes
Prometheus
ArgoCD
Service mesh

Job description

Plaid is seeking a Staff Site Reliability Engineer on Release Engineering to scale reliability across product engineering. You will architect the SLO and error-budget programs, drive progressive delivery, and ensure new products are production-ready.

By partnering with Platform and Infrastructure teams, you will translate complex production needs into intuitive, self-service tooling, lead incident responses, and shape deployment health for faster, safer software delivery.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer, Release Engineering
Staff Site Reliability Engineer, Release Engineering

Plaid Inc • New York (NY)

Hybrid
USD 130,000 - 180,000
Comprehensive benefit plan (medical, dental, vision, 401(k))
Equity and/or commission depending on position
Equal opportunity employer
Remote Release Engineer (SRE) for Scalable Reliability
Remote Release Engineer (SRE) for Scalable Reliability

United States Digital Space LLC • United States

Remote
USD 140,000 - 200,000
ESOP
Tech Allowance
Health Benefits
+3
SRE Release Engineer — Reliability for Global Deployments
SRE Release Engineer — Reliability for Global Deployments

Icehouseventures • United States

Remote
USD 140,000 - 190,000
Fully Remote
ESOP
Tech Allowance
+4
Senior SRE: Scale Reliability Leader (Hybrid)
Senior SRE: Scale Reliability Leader (Hybrid)

Plenful • San Francisco (CA)

Hybrid
USD 180,000 - 250,000
Healthcare Coverage
401(k) with Company Match
Equity
+5
Staff SRE & Platform Reliability Architect
Staff SRE & Platform Reliability Architect

Grailbio • Edison (CA)

On-site
USD 169,000 - 224,000
Flexible time-off
401(k) with employer match
Medical, dental, vision coverage
Staff SRE: Platform Reliability Architect
Staff SRE: Platform Reliability Architect

Anduril Industries • Costa Mesa (CA)

On-site
USD 191,000 - 253,000
Staff SRE — Platform Reliability Lead
Staff SRE — Platform Reliability Lead

United States Digital Space LLC • United States

Hybrid
USD 127,000 - 161,000
Deutschlandticket (Germany-wide public
28 vacation days
Work from abroad up to 10 days/year
+9
Senior SRE: Scale Reliability for Health AI Platform
Senior SRE: Scale Reliability for Health AI Platform

RXinsider LTD. • San Francisco (CA)

Hybrid
USD 150,000 - 230,000
Healthcare Coverage
401(k) Match
Equity
+5
Senior SRE – AI Cloud Reliability & Incidents
Senior SRE – AI Cloud Reliability & Incidents

PLAUD • United States

Hybrid
USD 140,000 - 190,000
ESOP
Hybrid work model
Health & retirement benefits
+4
Staff SRE: Reliability Architect for Enterprise Platforms
Staff SRE: Reliability Architect for Enterprise Platforms

Anduril Industries • Costa Mesa (CA)

On-site
USD 191,000 - 253,000