Senior SRE: GCP, Kubernetes, Terraform

Stord-Warehous

Atlanta (GA)

Remote

USD 150,000 - 210,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Stord is seeking an experienced Site Reliability Engineer to strengthen our infrastructure across GKE, Cloud Run, and AlloyDB. You’ll own Terraform modules, optimize containerized workloads, and lead incident response in a fast-growing environment.

You’ll collaborate with product and engineering teams to improve deployment practices, reduce toil, and raise reliability. A production mindset, strong communication, and adaptability to AI-assisted tooling are essential.

Qualifications

  • 5+ years in SRE, platform, or infrastructure engineering with ownership of complex systems
  • Strong hands-on experience with GCP core services (GKE, Cloud Run, AlloyDB, networking, IAM)
  • Fluent in Docker and Kubernetes with real workloads
  • Deep Terraform experience; reusable modules and state-aware work
  • Proficient in TypeScript, Python, or Go for tooling and automation
  • Proficient in building actionable monitoring and alerting (Datadog, Prometheus/Grafana)
  • Knowledge of distributed systems, failure modes, and scale
  • Experience with Git-based collaboration and incident management
  • Strong communication and collaboration, with a production-focused mindset
  • Familiarity with AI-assisted development tools is a plus

Responsibilities

  • Define and implement scalable, reliable infrastructure on GCP (GKE, Cloud Run, AlloyDB, networking)
  • Own IaC with Terraform: modules, policies, patterns
  • Manage container workloads on Kubernetes, sizing, performance
  • Reduce toil via automation and better defaults
  • Build and manage observability with Datadog (APM, logs, RUM)
  • Design and maintain CI/CD pipelines in GitHub Actions
  • Develop disaster recovery and business-continuity plans
  • Lead incident response and post-incident reviews
  • Collaborate across teams to improve deployment practices and reliability
  • Contribute to SRE and infra best practices and on-call rotations

Skills

GCP fundamentals
Kubernetes
Terraform
Docker
Programming: TS/Python/Go
Observability: Datadog
Distributed systems
Git workflows
Incident management
Communication
Ownership
AI tooling for coding
Cost engineering

Tools

Terraform
GKE

Job description

Stord is seeking an experienced Site Reliability Engineer to strengthen our infrastructure across GKE, Cloud Run, and AlloyDB. You’ll own Terraform modules, optimize containerized workloads, and lead incident response in a fast-growing environment.

You’ll collaborate with product and engineering teams to improve deployment practices, reduce toil, and raise reliability. A production mindset, strong communication, and adaptability to AI-assisted tooling are essential.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE: Cloud Platform Reliability & Automation
Senior SRE: Cloud Platform Reliability & Automation

Stord, Inc. • United States

Remote
USD 140,000 - 180,000
Senior SRE: Cloud Reliability, Terraform & Kubernetes (Remote)
Senior SRE: Cloud Reliability, Terraform & Kubernetes (Remote)

Motion Recruitment • Chicago (IL)

On-site
USD 140,000 - 190,000
Remote Senior SRE: AWS, Kubernetes & Terraform
Remote Senior SRE: AWS, Kubernetes & Terraform

Motion Recruitment Partners LLC • Chicago (IL), Northern (KY)

Hybrid
USD 140,000 - 170,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kovoro • Denver (CO), Northern (KY)

Hybrid
USD 150,000 - 190,000
Senior SRE: Enterprise-Scale Cloud, Terraform & Telemetry
Senior SRE: Enterprise-Scale Cloud, Terraform & Telemetry

Staffing Science • Arizona

On-site
USD 180,000 - 240,000
Senior Site Reliability Engineer — Remote, Cloud Infra Lead
Senior Site Reliability Engineer — Remote, Cloud Infra Lead

The Consensus • United States

On-site
USD 140,000 - 190,000
Senior SRE: AI Cloud Infra, Kubernetes & Terraform (Remote)
Senior SRE: AI Cloud Infra, Kubernetes & Terraform (Remote)

Motion Recruitment • Mount Laurel Township (NJ)

Remote
USD 140,000 - 190,000
Remote equipment stipend
Annual learning and development budget
Equity / Stock Options
+1
GCP Cloud Engineer | Terraform, Kubernetes & AI Ops
GCP Cloud Engineer | Terraform, Kubernetes & AI Ops

Stylitics • New York (NY)

On-site
USD 140,000 - 160,000
Vision and dental insurance fully paid
Medical plan
Stock options
+5
Senior SRE - Kubernetes, Observability & Automation (Remote)
Senior SRE - Kubernetes, Observability & Automation (Remote)

Camunda • Atlanta (GA)

Remote
USD 150,000 - 242,000
Remote work
Annual company events
Health & wellbeing
+2
Site Reliability Engineer
Site Reliability Engineer

NextGen | GTA: A Kelly Telecom Company • Mount Laurel Township (NJ)

On-site
USD 110,000 - 170,000