Staff SRE: Multi-Region AWS, Kubernetes & Observability

GrubHub Inc.

New York (NY)

On-site

USD 209,000 - 217,000

Full time

6 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity
401(k)
Health, dental, vision
Disability insurance
Paid time off
Parental leave
Discounted meals
Exclusive perks

Job summary

Wonder is seeking a Staff-level SRE/Infrastructure engineer in New York to architect resilient production services, manage AWS infrastructure, and own CI/CD pipelines. You will lead incident response and multi-region design supporting seasonal traffic growth.

Expertise in Terraform/Terraspace, EKS, and observability is required. Experience with Python/Go and PCI-compliant environments is highly valued. This role offers a competitive package and equity opportunities.

Qualifications

  • Demonstrated senior-level ownership of production platforms end-to-end.
  • Experience driving incident response and architecture reviews across teams.
  • 8+ years in SRE, DevOps or infrastructure engineering with broad impact.

Responsibilities

  • Architect resilient, self-healing systems and co-own design of production services.
  • Own multi-region resilience, failover readiness, runbooks and drills, and data-layer replication.
  • Manage AWS infrastructure as code with Terraform/Terraspace; Kubernetes platform (EKS) upgrades and autoscaling.
  • Own observability pipelines: logging, metrics, tracing and alerting; optimize signal quality and cost.
  • Drive reliability using SLOs and telemetry data; close observability gaps to prevent incidents.
  • Design scaling and capacity strategies for seasonal traffic peaks.
  • Own cloud cost accountability: right-sizing and reservations; track savings.
  • Build and maintain CI/CD pipelines and deployment tooling.
  • Operate within PCI-scoped environments; maintain segregated clusters and access paths.
  • Shape incident management: lead incident response, postmortems and architecture reviews.

Skills

Terraform
Terraspace
Kubernetes (EKS)
AWS multi-region
CI/CD
Python/Go
Datastores
RabbitMQ/SQS
Observability
PCI compliance
Cloud fundamentals
Monitoring & telemetry

Tools

Terraspace
GitHub Actions
Jenkins

Job description

Wonder is seeking a Staff-level SRE/Infrastructure engineer in New York to architect resilient production services, manage AWS infrastructure, and own CI/CD pipelines. You will lead incident response and multi-region design supporting seasonal traffic growth.

Expertise in Terraform/Terraspace, EKS, and observability is required. Experience with Python/Go and PCI-compliant environments is highly valued. This role offers a competitive package and equity opportunities.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Staff SRE: Multi-Region AWS, Kubernetes & CI/CD
Senior Staff SRE: Multi-Region AWS, Kubernetes & CI/CD

Tapingo Ltd • New York (NY)

On-site
USD 209,000 - 217,000
Equity
401(k)
Medical, dental, vision plans
+1
Senior SRE - Multi-Region AWS, Kubernetes & CI/CD
Senior SRE - Multi-Region AWS, Kubernetes & CI/CD

Grubhub Holdings Inc. • New York (NY)

On-site
USD 209,000 - 217,000
Equity
401(k)
Medical/Dental/Vision
Staff SRE: Multi-Region AWS, Kubernetes & CI/CD
Staff SRE: Multi-Region AWS, Kubernetes & CI/CD

Tapingo Ltd • United States

On-site
USD 209,000 - 217,000
Equity
401(k)
Medical/Dental/Vision
+3
Senior SRE: Multi-Region, High Availability
Senior SRE: Multi-Region, High Availability

Radar • New York (NY)

On-site
USD 200,000 - 300,000
Competitive salary
Stock options
401(k) match
+10
Remote SRE: Multi-Region Infra, SLOs & Observability
Remote SRE: Multi-Region Infra, SLOs & Observability

JumpCloud Inc. • United States

Remote
USD 150,000 - 210,000
Remote-first culture
Global teams
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kovoro • Denver (CO), Northern (KY)

On-site
USD 150,000 - 190,000
Staff SRE: Scale, Observability & Automation Leader
Staff SRE: Scale, Observability & Automation Leader

Replit • Northern (KY)

Hybrid
USD 180,000 - 260,000
Salary & equity
401(k) matching
Health, dental, vision, life
+9
Senior Site Reliability Engineer
Senior Site Reliability Engineer

The ReWork Group • New York (NY)

On-site
USD 120,000 - 160,000
Senior SRE: Enterprise-Scale Cloud, Terraform & Telemetry
Senior SRE: Enterprise-Scale Cloud, Terraform & Telemetry

Staffing Science • Arizona

On-site
USD 180,000 - 240,000
Remote Senior SRE: AWS, Kubernetes & Terraform
Remote Senior SRE: AWS, Kubernetes & Terraform

Motion Recruitment Partners LLC • Chicago (IL), Northern (KY)

Hybrid
USD 140,000 - 170,000