Site Reliability Manager

Google

Bengaluru

On-site

INR 3,000,000 - 6,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Google is seeking a Site Reliability Engineering (SRE) leader in Bengaluru to manage a team of SREs focused on Google’s enterprise services. You will drive roadmaps, design reviews, and capacity planning to ensure reliability, scalability and uptime for critical systems.

You will mentor engineers, improve service maturity, and partner with product and operations teams to deliver robust platforms with automation at scale.

Qualifications

  • Bachelor's degree in Computer Science or a related technical field or equivalent practical experience.
  • 5 years of experience building or managing distributed systems or cloud infrastructure, with a focus on Kubernetes.
  • 5 years of experience in people management.
  • Experience with site reliability engineering, system design, distributed computing.

Responsibilities

  • Manage a team of 6-10 site reliability engineers supporting Google’s enterprise services.
  • Develop roadmaps, planning, objectives and OKRs to move forward the maturity of the managed services; engage in lifecycle from inception to refinement.
  • Support services before live through design consulting, platform development, capacity planning and launch reviews; maintain services post-launch by monitoring availability and health.
  • Scale systems sustainably through automation and changes that improve reliability and velocity.
  • Practice sustainable incident response to meet service level objectives.

Skills

People management
Distributed systems
Cloud infrastructure
Kubernetes
Site Reliability Engineering
System design

Education

Bachelor's degree in Computer Science or related field

Tools

SAP
ERP systems
Enterprise tooling

Job description

Minimum qualifications
  • Bachelor's degree in Computer Science, a related technical field, or equivalent practical experience.
  • 5 years of experience building or managing distributed systems or cloud infrastructure, with a focus on Kubernetes.
  • 5 years of experience in people management.
  • Experience with site reliability engineering, system design, distributed computing.
Preferred qualifications
  • 5 years of experience in people management, with managing distributed, multi‑site teams through engineering managers or tech leads.
  • Experience in Enterprise tooling and technology.
  • Experience in Systems, Applications, and Products (SAP) or other Enterprise Resource Planning (ERP) systems.
About The Job

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault‑tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally‑visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE will keep an ever‑watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, you’ll have the opportunity to manage the complex challenges of scale which are unique to Google Cloud, while using your expertise in coding, algorithms, complexity analysis and large‑scale system design. SRE’s culture of intellectual curiosity, problem solving and openness is key to its success. Our organization brings together people with a wide variety of backgrounds, experiences and perspectives. We encourage them to collaborate, think big and take risks in a blame‑free environment. We promote self‑direction to work on meaningful projects, while we also strive to create an environment that provides the support and mentorship needed to learn and grow. Core Enterprise System (CES) SRE is part of Corporate Engineering‑Site Reliability Engineering (SRE). We provide SRE support to Enterprise applications within Google, powering key verticals such as Finance, Legal, Supply Chain, and HR. Our mission is to deliver service excellence with engineering, innovation and customer focus and transform Google's enterprise domain. Google is an engineering company at heart. We hire people with a broad set of technical skills who are ready to take on some of technology's greatest challenges and make an impact on users around the world. At Google, engineers not only revolutionize search, they routinely work on scalability and storage solutions, large‑scale applications and entirely new platforms for developers around the world. From Google Ads to Chrome, Android to YouTube, social to local, Google engineers are changing the world one technological achievement after another.

Responsibilities
  • Manage a team of 6‑10 site reliability engineers supporting Google’s enterprise services.
  • Develop roadmaps, planning, objectives and key results (OKRs) to move forward the maturity of the managed services. Engage in and improve the whole lifecycle of services from inception and design, through deployment, operation and refinement.
  • Support services before they go live through activities such as system design consulting, developing software platforms and frameworks, capacity planning and launch reviews. Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
  • Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Practice sustainable incident response ensuring services meet their service level objectives.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Manager
Site Reliability Manager

Google Inc. • Bengaluru

On-site
INR 4,000,000 - 8,000,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Zorba AI • Chennai District

On-site
INR 1,200,000 - 2,400,000
Intermediate Applications Developer
Intermediate Applications Developer

UPS • Chennai District

On-site
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

UST • Pune District

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineers - Google Cloud Platform GCP RedHat OpenShift Administration
Site Reliability Engineers - Google Cloud Platform GCP RedHat OpenShift Administration

UPS • Thiruvallur District

On-site
INR 1,200,000 - 1,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps INC. • Pune District

On-site
INR 1,400,000 - 1,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps • Maharashtra

On-site
INR 4,000,000 - 7,000,000
Site Reliability Engineering (SRE) Lead
Site Reliability Engineering (SRE) Lead

Sidglobal • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Senior Cloud Site Reliability Engineer
Senior Cloud Site Reliability Engineer

Augusta Infotech • Bengaluru

Hybrid
INR 1,500,000 - 2,500,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Lonvec Technologies Private Limited • Hyderabad

On-site
INR 3,000,000 - 5,000,000