Staff Site Reliability Engineer

Tapingo Ltd

United States

On-site

USD 209,000 - 217,000

Full time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Equity
401(k)
Medical/Dental/Vision
Disability insurance
Paid time off
Meal discounts

Job summary

Grubhub is seeking a senior SRE/DevOps engineer to architect and operate scalable, self-healing systems for a high-traffic US dining platform. You will own multi-region AWS architecture, Kubernetes (EKS), and the observability stack, while advancing CI/CD pipelines and incident response across the production toolkit.

The role requires deep expertise in IaC, container orchestration, and cloud-native design, with a strong track record of improving reliability in a PCI-compliant, high-volume

Qualifications

  • Staff-level scope: owned a production platform end to end and led incident response.
  • Typically 8+ years in SRE, DevOps or infra engineering; breadth of impact matters.

Responsibilities

  • Architect resilient, self-healing systems and multi-region designs.
  • Own AWS infrastructure as code from design to rollout.
  • Own the Kubernetes platform, including EKS lifecycle and autoscaling.
  • Develop and maintain CI/CD pipelines and deployment tooling.
  • Shape incident management, postmortems and architecture reviews.
  • Drive reliability with SLOs and telemetry data; optimize cloud costs.

Skills

Terraform/Terraspace
Kubernetes (EKS)
AWS (multi-region)
CI/CD tooling
Python/Go

Tools

Jenkins
GitHub Actions
Terraform/Terraspace
Helm
Helmfile
MongoDB/Atlas
Redis/ElastiCache
RabbitMQ/AmazonMQ
SQS

Job description

About The Opportunity

About The Opportunity This role is crucial for simplifying the dining experience for students across the US. You will be instrumental in architecting resilient and self-healing solutions, managing AWS infrastructure, closing observability gaps, designing scaling approaches, and shaping incident management processes. Your contributions will span the entire development lifecycle, encompassing the building and maintenance of CI/CD pipelines. Your role will be pivotal in ensuring the platform's scalability to support Grubhub's continuously expanding customer base, evidenced by the addition of 30 new campuses and a 25% year-over-year increase in order volume.

Impact You'll Make

Impact You'll Make Architect resilient, self-healing systems and co-own the design of critical production services. Own multi-region resilience: active-standby architecture, regional failover readiness, runbooks and drills, RTO/RPO targets, and the data-layer replication behind them (ElastiCache Global Datastore, MongoDB/Atlas, RDS). Own AWS infrastructure as code - Terraform/Terraspace, Helm and Helmfile - from design through rollout. Own the Kubernetes platform: EKS lifecycle, controller and add-on upgrades, ingress/gateway, and autoscaling (HPA/KEDA). Own the observability platform end to end - logging, metrics, tracing and alerting pipelines - including its signal quality and its cost. Drive reliability improvements using SLOs and telemetry data, closing observability gaps before they turn into incidents. Design scaling and capacity strategy against a strongly seasonal traffic profile that peaks at back-to-school. Own cloud cost accountability for the platform: right-sizing, reservations, and tracking realized savings. Build and maintain CI/CD pipelines and the deployment tooling the team depends on. Operate within PCI-scoped environments, respecting segregated clusters and access paths. Shape incident management: lead incident response, postmortems, failure analysis, service reviews and architecture review.

What You'll Bring To The Table

What You'll Bring To The Table Experience: Demonstrated Staff-level scope: has owned a production platform end to end, driven technical direction across teams without formal authority, and led both incident response and architecture review. Typically 8+ years in SRE, DevOps or infrastructure engineering, though scope and impact weigh more than tenure.

Technical Skills

Technical Skills: Deep experience with Infrastructure as Code - Terraform (with Terraspace or a similar wrapper) - owning modules, state and rollout across multiple environments. Kubernetes at operator depth: EKS lifecycle, Helm and Helmfile, controllers and add-ons, ingress/gateway, and autoscaling (HPA/KEDA). Multi-region AWS architecture, including failover design and the data-layer replication it depends on. Deep knowledge of CI/CD tools (e.g., Jenkins, GitHub Actions). Software engineering experience in Python, Go, or a similar object-oriented language. Proficiency with datastores (MySQL, MongoDB/Atlas, Redis/ElastiCache) and message brokers (RabbitMQ/AmazonMQ, SQS). Experience with Microservice Architecture and Application Design. Distributed monitoring experience, including SLOs, metrics, tracing and log pipelines. Strong working knowledge of cloud fundamentals (AWS compute/containers, storage, Linux, networking). Comfort operating in compliance-scoped environments (PCI or equivalent).

Soft Skills

Soft Skills: Strong technical writing, documentation, and communication skills. Experience with highly trafficked web-based services. Ability to set technical direction and build consensus across engineering teams in multiple time zones.

New York Base Salary

New York Base Salary: $208,500-$216,500. Wonder uses geographic-specific salary structures, which means the salary offered may vary depending on where the job is located. The final salary offer will take into account various factors, such as the candidate's skills, education, training, credentials, and experience.

Benefits

Benefits The benefits applicable to this role include a competitive compensation package with equity and a 401(k). We also offer a choice of medical, dental, and vision plans, company paid short and long term disability coverage, paid time off including flexible time off for exempt employees, paid vacation for non-exempt employees, and paid sick leave in compliance with applicable law in addition to paid parental leave, discounted meals and exclusive perks across the Wonder family of brands. Eligibility, effective dates, and available plan options vary by employment classification and location.

A Final Note

A Final Note At Wonder, we build the best teams by hiring with an objective lens — evaluating people for their potential while championing diversity, equity, and inclusion. We do not discriminate based on race, color, religion, gender identity or expression, sexual orientation, national origin, age, military service eligibility, veteran status, marital status, disability, or any other protected class. As part of our commitment to fair and compliant hiring practices, Wonder participates in the federal government's E-Verify program to confirm employment eligibility. If you need an accommodation during the interview process, please let your recruiter know.

Join Our Team

Join Our Team At Grubhub, we champion restaurants from coast to coast. Restaurants sit at the heart of communities. It’s our mission to strengthen their roots, deepen their connections, and increase the positive impact they have on people and society. Grubhub, part of Wonder, delivers the best local, authentic cuisine right to diners’ doors—and new customers and billions in revenue to local businesses. Featuring over hundreds of thousands of merchants in over 4,000 cities nationwide, our innovative technology, user-friendly platforms, and streamlined delivery capabilities have made us an industry leader in the world of online food ordering. Since we opened our doors in 2004, Grubhub has been opening doors all across the country. Bakery doors in Hyde Park, jibarito joint doors in Queens, and doors of opportunity all across the country. Join our team and help us open more.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Product Manager
Senior Product Manager

Grubhub Holdings Inc. • United States

On-site
USD 138,000 - 146,000
Equity & 401(k)
Health plans (medical, dental, vision)
Disability coverage
+4
Senior Product Manager
Senior Product Manager

GrubHub Inc. • New York (NY)

On-site
USD 138,000 - 146,000
Equity
401(k) plan
Medical, dental, and vision plans
+3
Associate Account Manager, Enterprise Restaurants
Associate Account Manager, Enterprise Restaurants

Grubhub Holdings Inc. • Chicago (IL)

Hybrid
USD 45,000 - 56,000
Equity
401(k)
Health plans
+4
Onboarding Associate, SMB
Onboarding Associate, SMB

Grubhub Holdings Inc. • Chicago (IL)

Hybrid
USD 42,000 - 52,000
Medical/Dental/Vision
401(k) plan
Discounted meals
+1
Senior Staff Machine Learning Engineer
Senior Staff Machine Learning Engineer

Grubhub Holdings Inc. • United States

On-site
USD 240,000 - 250,000
Equity
401(k)
Medical, dental, and vision plans
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grubhub Holdings Inc. • Chicago (IL)

On-site
USD 158,500 - 172,000
Procurement Associate
Procurement Associate

GrubHub Inc. • Northern (KY)

On-site
USD 75,000 - 79,000
Equity
401(k)
Medical/Dental/Vision
+3
Procurement Associate
Procurement Associate

Wonder Group, INC • Lake Parsippany (NJ)

On-site
USD 75,000 - 79,000
Equity
401(k)
Medical plan
+6
Senior Associate, Finance
Senior Associate, Finance

Wonder Group, INC • Chicago (IL)

On-site
USD 87,000 - 97,000
Equity
401(k)
Medical/dental/vision
+4
Business Development Representative (Atlanta)
Business Development Representative (Atlanta)

Wonder • Georgia

On-site
USD 42,000 - 105,000
Equity
401(k)
Medical/Dental/Vision
+2