Staff Site Reliability Engineer - Data Pipeline Resilience

Google Inc.

Pittsburgh, Northern (Allegheny County, KY)

Hybrid

USD 207,000 - 300,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity
Bonus target
Benefits

Job summary

Google Pittsburgh, PA, USA is seeking a Staff Site Reliability Engineer for the AQI Data SRE team. You will define technical goals, oversee architectural roadmaps, and drive reliability across SAGE/SPARK production stacks.

You’ll lead design reviews, simplify backends, and ensure production readiness for Ads launches. The role requires 8 years of software development, 3 years of leadership, and deep expertise in distributed systems and large-scale data processing.

Qualifications

  • 8 years of experience with software development in one or more programming languages.
  • 3 years of experience leading projects.
  • 3 years of experience designing, analyzing, and troubleshooting distributed systems.
  • Experience with large-scale data processing.

Responsibilities

  • Define the technical goal, architectural roadmap, and reliability strategy for AQI Data SRE across the SAGE/SPARK production stack.
  • Partner with dev leadership to lead system design reviews, drive backend simplification and resource isolation, and ensure production readiness for flagship Ads launches.
  • Architect and lead the implementation of critical infrastructure projects, including data-pipeline resilience, automated rollback/restart platforms, and consolidated observability.
  • Mentor and grow engineers on the team, and push for software engineering excellence, AI-first tooling, and SRE best practices; drive the operational maturity handover that helps partner dev teams mature.
  • Lead the response to complex production incidents, correlate production signals with customer and business impact, and drive high-impact post-incident architectural improvements and blameless postmortems and participate in a healthy tier 2 oncall rotation for core shared SAGE/SPARK infrastructure.

Skills

Distributed systems design
Large-scale data processing
Leadership/mentoring
Problem solving

Education

Bachelor's degree in Computer Science or related field
Master's degree in CS or Engineering (preferred)

Job description

Google Pittsburgh, PA, USA is seeking a Staff Site Reliability Engineer for the AQI Data SRE team. You will define technical goals, oversee architectural roadmaps, and drive reliability across SAGE/SPARK production stacks.

You’ll lead design reviews, simplify backends, and ensure production readiness for Ads launches. The role requires 8 years of software development, 3 years of leadership, and deep expertise in distributed systems and large-scale data processing.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: AI-Driven, Resilient Data Pipelines
Senior SRE: AI-Driven, Resilient Data Pipelines

Google • Pittsburgh

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Staff Site Reliability Engineer, AQI Data SRE
Staff Site Reliability Engineer, AQI Data SRE

Socket.dev • Pittsburgh

On-site
USD 207,000 - 300,000
Staff Site Reliability Engineer, AQI Data SRE
Staff Site Reliability Engineer, AQI Data SRE

Google • Pittsburgh

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Senior SRE - AI-Driven, Scalable Data Pipelines
Senior SRE - AI-Driven, Scalable Data Pipelines

Socket.dev • Pittsburgh

On-site
USD 207,000 - 300,000
Staff Site Reliability Engineer, AQI Data SRE
Staff Site Reliability Engineer, AQI Data SRE

Google Inc. • Pittsburgh, Northern (KY)

Hybrid
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Staff Site Reliability Engineer, AQI Data SRE
Staff Site Reliability Engineer, AQI Data SRE

Google Inc. • Pittsburgh, Northern (KY)

Hybrid
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Senior Site Reliability Engineer, Scalable Systems & Automation
Senior Site Reliability Engineer, Scalable Systems & Automation

Google Inc. • San Jose (CA)

On-site
USD 207,000 - 300,000
Equity
Benefits
Staff SRE: Scale, Automation & Resilient Systems
Staff SRE: Scale, Automation & Resilient Systems

Google Inc. • Pittsburgh

On-site
USD 207,000 - 301,000
20% bonus target
Equity
Comprehensive benefits
Senior Staff SRE: Scale, Reliability & Automation
Senior Staff SRE: Scale, Reliability & Automation

Google • New York (NY)

On-site
USD 262,000 - 364,000
SRE Engineering Manager - AI Foundry & Global Reliability
SRE Engineering Manager - AI Foundry & Global Reliability

Socket.dev • San Jose (CA)

On-site
USD 207,000 - 300,000