Senior SRE: AI-Driven, Resilient Data Pipelines

Google

Pittsburgh (Allegheny County)

On-site

USD 207,000 - 300,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Equity
Bonus target
Benefits

Job summary

Google is hiring for a Site Reliability Engineer to support AQI Data SRE across the SAGE/SPARK stack. The role focuses on reliability, scalability, and data-pipeline resilience in a large-scale ads infrastructure.

You will mentor engineers, drive AI-first tooling, and participate in blameless postmortems while maintaining a high standard of production readiness and collaboration with cross-functional teams.

Qualifications

  • Bachelor's degree in CS or related field or equivalent practical experience.
  • 8 years of software development experience.
  • 3 years leading projects.
  • 3 years designing, analysing, and troubleshooting distributed systems.
  • Experience with large-scale data processing.

Responsibilities

  • Define the technical goal, architectural roadmap, and reliability strategy for AQI Data SRE across the SAGE/SPARK production stack.
  • Partner with dev leadership to lead system design reviews, drive backend simplification and resource isolation, and ensure production readiness for flagship Ads launches.
  • Architect and lead the implementation of critical infrastructure projects, including data-pipeline resilience, automated rollback/restart platforms, and consolidated observability.
  • Mentor and grow engineers on the team, and push for software engineering excellence, AI-first tooling and SRE best practices; drive the operational maturity handover that helps partner dev teams mature.
  • Lead the response to complex production incidents, correlate production signals with customer and business impact, and drive high-impact post-incident architectural improvements and blameless postmortems and participate in a healthy tier 2 oncall rotation for core shared SAGE/SPARK infrastructure.

Skills

8 years software development
3 years leading projects
3 years distributed systems design/trb
Large-scale data processing
AI-assisted tooling / reliability
Cross-functional architectural impact
Collaboration with development teams
Blameless postmortems / SRE practices

Education

Bachelor's degree in Computer Science or related field or equivalent
Master's degree in Computer Science or Engineering

Tools

None

Job description

Google is hiring for a Site Reliability Engineer to support AQI Data SRE across the SAGE/SPARK stack. The role focuses on reliability, scalability, and data-pipeline resilience in a large-scale ads infrastructure.

You will mentor engineers, drive AI-first tooling, and participate in blameless postmortems while maintaining a high standard of production readiness and collaboration with cross-functional teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer - Data Pipeline Resilience
Staff Site Reliability Engineer - Data Pipeline Resilience

Google Inc. • Pittsburgh, Northern (KY)

Hybrid
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Senior SRE - AI-Driven, Scalable Data Pipelines
Senior SRE - AI-Driven, Scalable Data Pipelines

Socket.dev • Pittsburgh

On-site
USD 207,000 - 300,000
Staff Site Reliability Engineer, AQI Data SRE
Staff Site Reliability Engineer, AQI Data SRE

Socket.dev • Pittsburgh

On-site
USD 207,000 - 300,000
Staff Site Reliability Engineer, AQI Data SRE
Staff Site Reliability Engineer, AQI Data SRE

Google • Pittsburgh

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Staff Site Reliability Engineer, AQI Data SRE
Staff Site Reliability Engineer, AQI Data SRE

Google Inc. • Pittsburgh, Northern (KY)

Hybrid
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Senior SRE Lead: AViD Infra & AI Ops for Ads
Senior SRE Lead: AViD Infra & AI Ops for Ads

Google • Mountain View (CA)

On-site
USD 262,000 - 364,000
Senior SRE & Data Intelligence Engineering Manager
Senior SRE & Data Intelligence Engineering Manager

Google Inc. • San Jose (CA), Northern (KY)

Hybrid
USD 207,000 - 300,000
Equity
Benefits
Bonus target
SRE Engineering Manager - AI Foundry & Global Reliability
SRE Engineering Manager - AI Foundry & Global Reliability

Socket.dev • San Jose (CA)

On-site
USD 207,000 - 300,000
SRE Manager: AI Infrastructure & Data Intelligence
SRE Manager: AI Infrastructure & Data Intelligence

Socket.dev • San Jose (CA)

On-site
USD 207,000 - 300,000
Senior SRE: Data Intelligence & AI Platform
Senior SRE: Data Intelligence & AI Platform

Sabre • Southlake (TX)

On-site
USD 120,000 - 190,000
Competitive pay
Flexible work options
Comprehensive healthcare coverage
+4