Director, Site Reliability Engineering

Visa Hunt

Bengaluru

Hybrid

INR 3,000,000 - 6,000,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Okta seeks a visionary Director of Site Reliability Engineering to lead the India-based SRE organization, overseeing production services, databases, edge networking, K8s, CI/CD, observability, and FinOps. You will partner with global teams to deliver scalable, secure platforms and drive automation and IaC adoption across the org.

The role emphasizes leadership, strategic planning, incident management, and mentoring top SRE talent to support Okta's high-availability SaaS services in a hybrid work

Qualifications

  • 16+ years in site reliability, infrastructure, or production engineering roles.
  • 8+ years in technical leadership & people management.
  • Experience building or scaling offshore SRE teams and partnering with global counterparts.
  • 4+ years leading SRE for SaaS/Cloud in a public Cloud (AWS preferred).

Responsibilities

  • Build and lead a India-based SRE organization supporting Okta’s production fleet.
  • Partner with global engineering, product, and infrastructure leaders to deliver resilient services.
  • Define and execute the India SRE strategy in alignment with global reliability goals.
  • Lead post-incident reviews and drive root-cause analysis; participate in incident management and RCAs.
  • Implement automation and observability to reduce toil and improve efficiency; promote IaC, Kubernetes, and AI tooling.
  • Hire, mentor, and develop SRE talent across India; foster a reliability-focused engineering culture.

Skills

SRE leadership
Automation
Observability
Cloud-native architecture
Kubernetes
Terraform
CI/CD
Incident response
Cross-cultural collaboration
People management

Education

Computer Science Degree

Tools

AWS
Kubernetes
Terraform
CI/CD tooling

Job description

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era.

This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.

Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput, and 99.999 availability. We're looking for a technical leader to help us to continue to scale the service with great people and reliable, cost-effective and efficient infrastructure, processes and tooling.

As the Director of Site Reliability Engineering you will oversee the SRE organization focused on Okta platform, Databases, Edge networking, K8s platform, CI/CD, Observability, FinOps, and automation platform & tooling.

Job Duties and Responsibilities:
  • Build and lead a high-caliber India-based SRE organization supporting Okta’s production fleet.
  • Partner with global engineering, product, and infrastructure leaders to deliver resilient, scalable, and secure services.
  • Define and execute the India SRE strategy in alignment with global reliability goals.
  • Lead post-incident reviews, drive root-cause analysis, and ensure long-term corrective actions. Participate in incident management, on-call rotations, and blameless RCAs.
  • Implement automation and observability to reduce manual toil and improve operational efficiency. Drive adoption of modern infrastructure practices: infrastructure as code (Terraform), container orchestration (Kubernetes), and AI within Infrastructure org.
  • Hire, mentor, and develop top SRE talent across India; build a strong engineering culture centered on reliability and innovation.
  • Foster collaboration across time zones with U.S. and EMEA teams.
  • Promote continuous learning, knowledge sharing, and process improvement.
  • Maintain a deep knowledge of industry best practices, evolving trends, and technologies.
  • Manage service and business expectations and prioritize resource allocation.
  • Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and intuitive self-service capabilities.
  • Improve SDLC processes for Cloud infrastructure as a code, including the maturity of CI/CD pipelines, change and release management.
Required Knowledge, Skills, and Abilities:
  • 16+ years of experience in site reliability, infrastructure, or production engineering roles.
  • 8+ years of experience in technical leadership & people management including managing managers. Strong expertise in automation, observability, performance optimization, and incident response.
  • Excellent leadership, communication, and cross-cultural collaboration skills—particularly in a global matrixed setup.
  • Experience building or scaling offshore SRE teams that partner with global counterparts.
  • 4+ years of experience running the SRE org supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus.
  • Strong expertise in cloud-native architectures, containerization (Kubernetes), IaC (Terraform), and CI/CD pipelines.
  • Demonstrated ability to lead cross-functional teams and manage large-scale programs
  • Effective verbal, written communication and interpersonal skills
Education and Training:

Computer Science Degree or related degree or equivalent experience

P23515_3244736

#LI-Hybrid

The Okta Experience
  • Supporting Your Well-Being
  • Driving Social Impact
  • Developing Talent and Fostering Connection + Community

We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.

Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.

If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.

Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Director, Site Reliability Engineering
Director, Site Reliability Engineering

Okta • Bengaluru

On-site
INR 4,500,000 - 8,000,000
Manager- Site Reliability Engineering
Manager- Site Reliability Engineering

Okta • Bengaluru

On-site
INR 3,000,000 - 6,000,000
Driving social impact
Talent development & community
Director, Site Reliability Engineering
Director, Site Reliability Engineering

United States Digital Space LLC • Bengaluru

Hybrid
INR 3,500,000 - 7,000,000
Staff Site Reliability Engineer - Ecosystem
Staff Site Reliability Engineer - Ecosystem

Triwill Group • Bengaluru

Hybrid
INR 3,500,000 - 5,500,000
Staff SRE for Cloud Network Infrastructure Team (Managing edge in multi-cloud(AWS & GCP), mTLS,[...]
Staff SRE for Cloud Network Infrastructure Team (Managing edge in multi-cloud(AWS & GCP), mTLS,[...]

Okta • Bengaluru

Hybrid
INR 3,500,000 - 6,000,000
Staff Site Reliability Engineer - Network Infrastructure
Staff Site Reliability Engineer - Network Infrastructure

Triwill Group • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Okta • Bengaluru

On-site
INR 2,000,000 - 3,000,000
Immersive onboarding experience
Equal Opportunity Employer benefits
Staff Site Reliability Engineer - (Infra)
Staff Site Reliability Engineer - (Infra)

Okta • Bengaluru

Hybrid
INR 1,500,000 - 2,000,000
Equal Opportunity Employer
Reasonable accommodation available
Collaboration with talented teams
Staff Site Reliability Engineer - Core iDaaS
Staff Site Reliability Engineer - Core iDaaS

Okta • Bengaluru

On-site
INR 6,000,000 - 9,000,000
Staff Site Reliability Engineer - Ecosystem
Staff Site Reliability Engineer - Ecosystem

Okta • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Immersive onboarding
Global community
Equal opportunity employer