Director, Site Reliability Engineering (Production Engineering)

Zscaler

Bengaluru

Hybrid

INR 6,000,000 - 12,000,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Time off plans
Parental leave
Retirement options
In-office perks

Job summary

Zscaler is seeking a Director of Site Reliability Engineering (Production Engineering) in a Hybrid (Bangalore) setup. You will lead SRE teams to achieve 99.99%+ availability across a cloud platform handling massive daily transactions, owning architecture, incident triage, and automation while growing engineering leadership.

Strong collaboration with India and global HQ is expected. You will drive operational rigor, implement robust alerting, and own capacity planning and strategy within a

Qualifications

  • AI/ML familiarity to optimize outcomes in the functional domain.
  • 12+ years engineering experience with 5+ years managing engineering managers.
  • Experience leading globally distributed engineering teams with US/intl stakeholders.
  • Deep understanding of observability architectures and telemetry pipelines.
  • Hands-on with Kubernetes and multi-cloud infrastructure.

Responsibilities

  • End-to-end ownership of availability, latency, and performance for Tier-0 and Tier-1 platforms.
  • Enforce multi-nines SLAs/SLOs and drive MTTR/MTTD improvements.
  • Oversee capacity planning and resource forecasting for scalable infra.
  • Implement high-signal alerting frameworks and runbooks to reduce blind spots.
  • Lead Production Engineering across India and align with global HQ.

Education

Bachelor's or Master's in CS/Software Engineering

Tools

Kubernetes
OpenTelemetry
Prometheus
Kafka
ClickHouse
Elasticsearch

Job description

Director, Site Reliability Engineering (Production Engineering)

Zscaler (NASDAQ: ZS) accelerates digital transformation so customers can be more agile, efficient, resilient, and secure. The Zscaler Zero Trust Exchange platform protects thousands of customers from cyberattacks and data loss by securely connecting users, devices, and applications in any location. Distributed across 160+ public exchanges globally and thousands of private exchanges at the edge, the SASE-based Zero Trust Exchange is the world’s largest in-line cloud security platform.

We believe the future of work is Human + AI and are building an AI-native enterprise where human potential is amplified by machine intelligence to solve the world’s hardest security challenges. Driven by deep customer obsession, we are committed to the mission, outcome, and to each other. We bring these commitments to life through three core behaviors: ownership and collaboration, trust through outcomes and impact, and a challenge culture with ongoing feedback. Ready to make an impact at the company pioneering security transformation in the AI era? Join us at Zscaler.

Role

We are looking for a Director, Site Reliability Engineering to join our team. This is a Hybrid (Bangalore) role, reporting to the VP, Engineering in the Cloud Infrastructure & Operations department.

In this role, you will lead our SRE teams in driving 99.99%+ availability across a cloud platform processing hundreds of billions of daily transactions. You will own system architecture, automated incident mitigation, and 24/7 operational reliability while scaling top engineering talent and aligning teams across India and global HQ.

What you’ll do (Role Expectations)

  • Take end-to-end operational accountability for the availability, latency, and performance of Tier-0 and Tier-1 internally facing platforms and backend infrastructure
  • Enforce multi-nines SLAs, SLOs, and error budgets across services while driving engineering initiatives to measurably reduce MTTD, MTTA, and MTTR
  • Oversee capacity planning, resource forecasting, and efficiency efforts to ensure key infrastructure scales reliably
  • Implement high-signal alerting frameworks and actionable runbooks, minimizing alert fatigue while ensuring zero unmonitored blind spots
  • Lead Production Engineering across India, driving operational rigor while coaching technical leadership on strategy, decision-making, and organizational scaling

Who You Are (Success Profile)

  • You act like an owner with a strong bias for action, operating with integrity, caring genuinely about outcomes, and navigating seamlessly between high-level strategy and hands-on execution.
  • You are a high-trust collaborator who is ambitious for the team, embracing a challenge culture through ongoing, respectful feedback to build trust and accelerate shared success.
  • You are customer-obsessed, building deep empathy for internal and external customers to anchor decisions in solving real-world problems and championing their needs from start to finish.
  • You are driven by innovation and energized by solving complex technical challenges, bringing deep curiosity and a constant search for better, more secure, and scalable solutions.
  • You lead with integrity, holding yourself and others to high standards of accountability while matching your words with consistent, transparent action.

What We’re Looking for (Minimum Qualifications)

  • Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain
  • 12+ years of engineering experience, including 5+ years managing engineering managers and distributed teams, with a proven track record as a site leader scaling and inspiring teams
  • Proven track record of operating effectively within globally distributed engineering teams, partnering closely with US-based leadership and stakeholders
  • Deep understanding of modern observability architectures—including distributed tracing, high-cardinality metrics engines, distributed log indexing, and streaming telemetry pipelines (e.g., Kafka, OpenTelemetry, Prometheus, ClickHouse, Elasticsearch)
  • Hands-on background architecting, deploying, and supporting hyper-scale cloud-native platforms running on Kubernetes and multi-cloud infrastructure
  • Proven expertise in site reliability engineering principles, chaos testing, capacity planning, and automated incident triage

What Will Make You Stand Out (Preferred Qualifications)

  • Proven ability to leverage AI technologies and workflows to drive measurable operational efficiency
  • Bachelor’s or Master’s degree in Computer Science, Software Engineering, or equivalent practical experience
  • Familiarity with high-volume data ingestion, backpressure management, real-time aggregation, and data lifecycle management
  • Executive-level communicator adept at translating complex technical trade-offs for leadership, with high EQ, resilience under pressure, and a passion for developing people and engineering culture

#LI-SK3

#LI-HYBRID

At Zscaler, we are committed to building a team that reflects the communities we serve and the customers we work with. We foster an inclusive environment that values all backgrounds and perspectives, emphasizing collaboration and belonging. Join us in our mission to make doing business seamless and secure.

Our Benefits program is one of the most important ways we support our employees. Zscaler proudly offers comprehensive and inclusive benefits to meet the diverse needs of our employees and their families throughout their life stages, including:

  • Time off plans for vacation and sick time
  • Parental leave options
  • Retirement options
  • In-office perks, and more!

Learn more about Zscaler's hybrid working model and benefits here .

By applying for this role, you adhere to applicable laws, regulations, and Zscaler policies, including those related to security and privacy standards and guidelines.

Zscaler is committed to providing equal employment opportunities to all individuals. We strive to create a workplace where employees are treated with respect and have the chance to succeed. All qualified applicants will be considered for employment without regard to race, color, religion, sex (including pregnancy or related medical conditions), age, national origin, sexual orientation, gender identity or expression, genetic information, disability status, protected veteran status, or any other characteristic protected by federal, state, or local laws. See more information by clicking on the Know Your Rights: Workplace Discrimination is Illegal link.

Pay Transparency

Zscaler complies with all applicable federal, state, and local pay transparency rules.

Zscaler is committed to providing reasonable support (called accommodations or adjustments) in our recruiting processes for candidates who are differently abled, have long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support.

Voluntary Self Identification

At Zscaler, we value diversity, equity, inclusion and belonging. We invite you to voluntarily respond to the question(s) below to help us measure the effectiveness of our outreach and recruitment. Responding is entirely voluntary and will not impact your application process. All responses will be kept confidential and handled in accordance with applicable privacy laws. Thank you for helping us create a more inclusive workplace.

Sex * Select...

By checking this box, I consent to Zscaler collecting, storing, and processing my responses to the demographic data surveys above. *

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Software Development Engineer
Sr. Software Development Engineer

Zscaler, Inc. • Bengaluru

Hybrid
INR 2,500,000 - 4,000,000
Hybrid work model
Time off & parental leave
Retirement options
+1
Staff Software Development Engineer (Backend - Java/API)
Staff Software Development Engineer (Backend - Java/API)

Zscaler, Inc. • Mohali

Hybrid
INR 4,000,000 - 7,000,000
Time off plans for vacation and sick
Parental leave options
Retirement options
+1
Staff Software Development Engineer - DevOps
Staff Software Development Engineer - DevOps

Zscaler • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Staff Site Reliability Engineer (Linux/Network troubleshooting/Scripting)
Staff Site Reliability Engineer (Linux/Network troubleshooting/Scripting)

Zscaler • Bengaluru

On-site
INR 3,500,000 - 6,000,000
Sr. Software Development Engineer (Backend Dev - C/C++)
Sr. Software Development Engineer (Backend Dev - C/C++)

Zscaler • Mohali

On-site
INR 2,500,000 - 4,500,000
Time off plans
Parental leave options
Retirement options
+1
Staff Software Development Engineer - Java/Go + Distributed Systems
Staff Software Development Engineer - Java/Go + Distributed Systems

Xapply • Bengaluru

Hybrid
INR 4,000,000 - 6,000,000
Health plans
Parental leave options
Retirement options
+2
Director, Professional Services & Consulting
Director, Professional Services & Consulting

Zscaler • Bengaluru

Hybrid
INR 6,000,000 - 10,000,000
Time off plans
Parental leave options
Retirement options
+1
Staff Software Development Engineer - Java/Go + Distributed Systems
Staff Software Development Engineer - Java/Go + Distributed Systems

Zscaler, Inc. • Bengaluru

Hybrid
INR 3,000,000 - 5,400,000
Time off plans for vacation and sick
Parental leave options
Retirement options
+1
Senior Staff Site Reliability Engineer
Senior Staff Site Reliability Engineer

Zscaler Softech • Bengaluru

Hybrid
INR 4,000,000 - 6,000,000
Health plans
Vacation & sick time
Parental leave
+3
Sr. Staff Site Reliability Engineer
Sr. Staff Site Reliability Engineer

Zscaler • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Health plans
Time off plans (vacation & sick)
Parental leave options
+3