Senior Director, Global Infra & Sovereign SRE

Google

Sunnyvale (CA)

On-site

USD 364,000 - 505,000

Full time

33 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Google’s Site Reliability Engineering (SRE) organization is hiring a senior leader in Technical Infrastructure to design, build, and run large-scale, fault-tolerant production systems. You will guide cross-functional teams across software and systems engineering to ensure reliability, capacity planning, and performance while driving automation and efficiency.

The role emphasizes leading distributed teams, growing and shaping engineering organizations, and delivering scalable platforms for

Qualifications

  • Bachelor’s degree in Computer Science or related field or equivalent practical experience.
  • 20 years of experience with system design, algorithms, data structures, analysis, and software design.
  • 15 years of experience managing a distributed team of engineers.
  • Experience growing and building teams.

Responsibilities

  • Create, design & lead innovative programs, software, & analytics that drive improvements to the reliability, scalability, efficiency, & velocity of new platform introductions & existing hardware, out machine management systems, & our node software.
  • Work in close partnership with product leads across multiple areas, on technologies up & down Google's tall & deep stack to design & help build fast, reliable & durable production systems.
  • Manage reliability for a massive, globally distributed "fleet of fleets." This includes managing lifecycle operations (Day 0 turnup, Day 1 install, Day 2 repairs) in addition to software reliability.
  • Own, innovate and create programs, software solutions, process innovations, and analytics that drive improvements to the availability, scalability, latency, and efficiency of Google’s technical Infrastructure and Google Cloud products.
  • Architect observability frameworks that provide operational health signals back to Google or empower air-gapped operators without leaking sovereign data.

Skills

System design
Algorithms
Data structures
Software design
Leadership

Education

Bachelor’s degree in Computer Science
Master’s degree or PhD in Computer Science

Job description

Google’s Site Reliability Engineering (SRE) organization is hiring a senior leader in Technical Infrastructure to design, build, and run large-scale, fault-tolerant production systems. You will guide cross-functional teams across software and systems engineering to ensure reliability, capacity planning, and performance while driving automation and efficiency.

The role emphasizes leading distributed teams, growing and shaping engineering organizations, and delivering scalable platforms for

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Director, Global Infrastructure & Sovereign SRE
Senior Director, Global Infrastructure & Sovereign SRE

Socket.dev • Sunnyvale (CA)

On-site
USD 364,000 - 505,000
Senior Director, Infrastructure and Sovereign Environments, SRE, Google Cloud
Senior Director, Infrastructure and Sovereign Environments, SRE, Google Cloud

Socket.dev • Sunnyvale (CA)

On-site
USD 364,000 - 505,000
Director, Cloud SRE & Sovereign Environments
Director, Cloud SRE & Sovereign Environments

Socket.dev • Sunnyvale (CA)

On-site
USD 307,000 - 427,000
Senior SRE Engineering Manager – Lead Uptime
Senior SRE Engineering Manager – Lead Uptime

Google • San Bruno (CA)

On-site
USD 262,000 - 364,000
Senior Director, Infrastructure and Sovereign Environments, SRE, Google Cloud
Senior Director, Infrastructure and Sovereign Environments, SRE, Google Cloud

Google • Sunnyvale (CA)

On-site
USD 364,000 - 505,000
Senior Staff SRE: Scale, Reliability & Automation
Senior Staff SRE: Scale, Reliability & Automation

Google • New York (NY)

On-site
USD 262,000 - 364,000
Senior Site Reliability Engineer, Scalable Systems & Automation
Senior Site Reliability Engineer, Scalable Systems & Automation

Google Inc. • San Jose (CA)

On-site
USD 207,000 - 300,000
Equity
Benefits
Director, Engineering, Sovereign Environments, SRE, Cloud
Director, Engineering, Sovereign Environments, SRE, Cloud

Socket.dev • Sunnyvale (CA)

On-site
USD 307,000 - 427,000
SRE Engineering Manager - AI Foundry & Global Reliability
SRE Engineering Manager - AI Foundry & Global Reliability

Socket.dev • San Jose (CA)

On-site
USD 207,000 - 300,000
Staff SRE Engineer — Home IoT & Cloud Reliability
Staff SRE Engineer — Home IoT & Cloud Reliability

Google • San Francisco (CA)

On-site
USD 207,000 - 300,000