Senior Site Reliability Engineer (Remote) – Scale & Reliability

WellSaid

United States

Remote

USD 140,000 - 190,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Stock options
Medical, dental, and vision insurance
401(k) plan matching
Generous vacation/PTO
Parental leave
Learning & development stipend
Home office stipend

Job summary

WellSaid Labs, a leading AI voiceover studio, seeks a Site Reliability Engineer to run production platforms and scale data pipelines. You will monitor, secure, and improve availability across services, lead incident response, and enable fast, reliable deployment across teams.

You should bring 5+ years in SRE/DevOps, strong cloud, Kubernetes, Docker, and IaC experience, plus a passion for scalable, reliable systems at scale.

Qualifications

  • 5+ years with a modern programming language (Golang, Typescript, Python, etc…).
  • Strong understanding of Infrastructure as Code (Terraform, Tofu, Pulumi).
  • Experience with GitOps/continuous delivery tooling (ArgoCD, Spacelift, Terraform Cloud).
  • Experience building and troubleshooting Kubernetes environments.
  • Ability to debug and solve issues in complex production environments.
  • Understanding of profiling applications and databases to identify performance issues.
  • Fluency in a UNIX shell for log analysis and operational tasks.
  • Experience with monitoring tools such as Grafana and Prometheus.

Responsibilities

  • Build and run our monitoring, tracing and alerting infrastructure.
  • Drive platform security initiatives to ensure system reliability.
  • Lead incident response and recovery including root cause analysis.
  • Run our platform with high availability and redundancy front and center.
  • Respond to alerts and be part of the on-call response team.
  • Improve deployment processes for faster, safer code changes.
  • Deliver a stable, scalable product platform enabling fast engineering delivery.
  • Identify novel ways to handle load and scale resource-intensive apps.
  • Help engineers write reliable code while shipping quickly.

Skills

Algorithms
Data structures
Systems design
Unix shell
Golang
Typescript
Python
Cloud (AWS/GCP)

Tools

Kubernetes
Docker
Terraform
Pulumi
ArgoCD
Spacelift
Terraform Cloud
Grafana
Prometheus

Job description

WellSaid Labs, a leading AI voiceover studio, seeks a Site Reliability Engineer to run production platforms and scale data pipelines. You will monitor, secure, and improve availability across services, lead incident response, and enable fast, reliable deployment across teams.

You should bring 5+ years in SRE/DevOps, strong cloud, Kubernetes, Docker, and IaC experience, plus a passion for scalable, reliable systems at scale.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE & Security Engineer — Scale & Reliability
Senior SRE & Security Engineer — Scale & Reliability

Remote Genie • Northern (KY)

Hybrid
USD 140,000 - 180,000
Competitive salary and stock options
Full medical, dental, and vision
Matching 401(k) plan
+3
Remote Senior SRE: Reliability, Security & Scale
Remote Senior SRE: Reliability, Security & Scale

WellSaid Labs, Inc. • Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive salary
Stock options
Medical, dental, and vision insurance
+4
Senior Software Engineer, Site Reliability & Security New Remote - United States
Senior Software Engineer, Site Reliability & Security New Remote - United States

WellSaid Labs, Inc. • Northern (KY)

Remote
USD 140,000 - 190,000
Competitive salary
Stock options
Medical, dental, and vision insurance
+4
Senior Site Reliability Engineer — AI Platform Scale
Senior Site Reliability Engineer — AI Platform Scale

Future Secure AI • Austin (TX)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer - AI-Driven Infra (Remote)
Senior Site Reliability Engineer - AI-Driven Infra (Remote)

ExpertVoice • United States

Remote
USD 170,000 - 185,000
Medical insurance
Dental insurance
Vision insurance
+4
Remote SRE: AI Platform Reliability & Automation
Remote SRE: AI Platform Reliability & Automation

Runpod • United States

On-site
USD 150,000 - 200,000
Remote work first
Competitive base salary
Stock options equity
+2
Senior Site Reliability Engineer — Remote Production Reliability
Senior Site Reliability Engineer — Remote Production Reliability

Fingerprint • Chicago (IL)

Remote
USD 152,000 - 205,000
Remote Senior SRE: Build Reliable, Scalable AI Infra
Remote Senior SRE: Build Reliable, Scalable AI Infra

Runware • Town of Sweden (NY)

On-site
USD 140,000 - 190,000
Generous paid time off
Meaningful stock options
Remote-first setup
+3
Senior Software Engineer, Site Reliability & Security WellSaid 9h ago
Senior Software Engineer, Site Reliability & Security WellSaid 9h ago

Remote Genie • Northern (KY)

On-site
USD 140,000 - 180,000
Competitive salary and stock options
Full medical, dental, and vision
Matching 401(k) plan
+3
Remote Site Reliability Engineer — Scale & Observability
Remote Site Reliability Engineer — Scale & Observability

Orkes • United States

Remote
USD 180,000 - 250,000
Comprehensive health coverage
Flexible PTO
Support for personal development