Remote SRE — Scale Real-World Freight Platform

Outpost

United States

Remote

USD 140,000 - 200,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Outpost is seeking a Site Reliability Engineer to own uptime and incident response as our platform scales. You will strengthen monitoring, automate remediation, and partner with engineering to triage alerts.

The role focuses on GCP infrastructure, database performance, and ML training reliability, with blameless postmortems and on-call participation. You’ll work with a small, high-conviction team building real software for live logistics yards, ensuring reliability as load grows and automation

Qualifications

  • 4+ years in an SRE or backend role with production on-call ownership.
  • Deep experience with a major cloud provider (GCP preferred) including compute, managed databases, object storage, networking.
  • Experience building monitoring/alerting/observability stacks (Grafana, Prometheus, Zabbix, Datadog, or similar).
  • Strong scripting/automation skills (Python, Bash, or similar).
  • Comfortable with containerized workloads (Docker) and CI/CD pipelines.
  • Track record of reducing incident volume or improving reliability metrics.

Responsibilities

  • Own reliability targets across backend/API, worker services, applications and CV pipeline; MTTR and root-cause follow-up.
  • Level up monitoring/alerting and build auto-remediation to scale on-call load.
  • Partner with engineering to build agents that triage alerts and handle routine remediation.
  • Harden and optimize GCP infrastructure (Cloud Run, Cloud SQL, GCS) for cost and performance as load scales.
  • Own database scale and performance; tuning, read replicas, and capacity planning.
  • Improve reliability of ML training and monitoring infrastructure in collaboration with CV/ML teams.
  • Run blameless postmortems and drive fixes for root causes; participate in on-call rotation.

Job description

Outpost is seeking a Site Reliability Engineer to own uptime and incident response as our platform scales. You will strengthen monitoring, automate remediation, and partner with engineering to triage alerts.

The role focuses on GCP infrastructure, database performance, and ML training reliability, with blameless postmortems and on-call participation. You’ll work with a small, high-conviction team building real software for live logistics yards, ensuring reliability as load grows and automation

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Outpost • United States

Remote
USD 140,000 - 200,000
Remote SRE: AI Platform Reliability & Automation
Remote SRE: AI Platform Reliability & Automation

Runpod • United States

On-site
USD 150,000 - 200,000
Remote work first
Competitive base salary
Stock options equity
+2
Remote Site Reliability Engineer II — Scale & Reliability
Remote Site Reliability Engineer II — Scale & Reliability

Mrsool • United States

Remote
USD 120,000 - 180,000
Remote work options
Competitive compensation
Learning stipend
Senior SRE: Scale Real-Time Platform with Automation
Senior SRE: Scale Real-Time Platform with Automation

Alembic • Dunwoody (GA)

On-site
USD 150,000 - 190,000
Site Reliability Engineer II — Scale & Resilience
Site Reliability Engineer II — Scale & Resilience

DAT Freight Solutions • Seattle (WA)

Hybrid
USD 95,000 - 134,000
Medical, Dental, Vision, Life, AD&D
401k matching
Employee Stock Purchase Plan
+4
SRE II: Scale Systems with Automation & Observability
SRE II: Scale Systems with Automation & Observability

DAT Freight & Analytics • Seattle (WA)

Hybrid
USD 95,000 - 134,000
Medical insurance
Dental insurance
Vision insurance
+5
Site Reliability Engineer II — Scale & Resilience
Site Reliability Engineer II — Scale & Resilience

DAT Freight Solutions • Portland (OR)

Hybrid
USD 95,000 - 134,000
Medical, Dental, Vision
401k matching
Flexible vacation
+1
Senior SRE: Observability, Automation & Scalable Systems
Senior SRE: Observability, Automation & Scalable Systems

Replit • Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive Salary & Equity
401(k) 4% match (US)
Health, Dental, Vision & Life
+7
SRE Platform Engineer II: Scale, Automate Cloud Systems
SRE Platform Engineer II: Scale, Automate Cloud Systems

DAT Freight & Analytics • Portland (OR)

Hybrid
USD 95,000 - 134,000
Medical, Dental, Vision
Parental Leave
Flexible Vacation Time
+2
Hybrid SRE Engineer for Scalable Reliability & Equity
Hybrid SRE Engineer for Scalable Reliability & Equity

EarnIn • Mountain View (CA)

Hybrid
USD 139,000 - 232,000