Director – Site Reliability Engineering (SRE)

Umanist NA

Hyderabad

On-site

INR 3,500,000 - 5,200,000

Full time

2 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Umanist NA in Hyderabad seeks a Director of Site Reliability Engineering (SRE) to lead reliability across multiple B2B SaaS products. The role combines software engineering fundamentals with deep SRE expertise in cloud infrastructure, observability, incident management, and CI/CD.

You will drive scalable reliability programs, establish SLI/SLO practices, and improve availability and performance. The leader will define and manage reliability KPIs, mentor engineering leaders, and collaborate with

Qualifications

  • Bachelor's degree and 18+ years of experience in Software Engineering, SRE, or related roles.
  • 5+ years of leadership experience at Director level or equivalent.
  • Strong AWS expertise and experience with cloud-based SaaS platforms.

Responsibilities

  • Lead and develop SRE teams supporting multiple SaaS products.
  • Establish and drive reliability strategy across the organization.
  • Define SLIs, SLOs, SLAs, error budgets, and reliability KPIs.
  • Drive observability, monitoring, alerting, and CI/CD improvements.
  • Architect and support applications and infrastructure on cloud platforms.
  • Mentor engineering leaders and senior SRE engineers.
  • Communicate reliability strategy and KPIs to senior leadership.

Skills

Kubernetes
Distributed systems
Metrics/KPIs
Container orchestration
Production operations
CI/CD
SRE leadership
AWS
Observability
SLO/SLI management

Education

Bachelor's degree in Computer Science, Information Systems, Engineering, or a related field

Tools

Terraform
Prometheus / Grafana
Datadog / New Relic / Splunk
Kubernetes
AWS CloudWatch

Job description

Job Title: Director - Site Reliability Engineering (SRE)

Industry: B2B SaaS / Cloud Product

Experience: 18+ Years

Leadership: 5+ Years at Director / Senior Engineering Leadership Level

Notice Period: Immediate to 30 Days

Job Location: Hyderabad

Role Overview

We are looking for an experienced Director of Site Reliability Engineering (SRE) to lead reliability and operational excellence across multiple SaaS products.

The ideal candidate combines strong software engineering fundamentals with deep expertise in SRE, cloud infrastructure, observability, monitoring, incident management, CI/CD, and distributed systems.

This leader will be responsible for building scalable reliability programs, improving availability and performance, establishing SLI/SLO practices, and driving operational excellence through measurable metrics and KPIs.

Key Responsibilities
  • Lead and develop SRE / Reliability Engineering teams supporting multiple SaaS products.
  • Establish and drive reliability engineering strategy across the organization.
  • Define and manage SLIs, SLOs, SLAs, error budgets, and reliability KPIs.
  • Drive observability, monitoring, alerting, and proactive performance management.
  • Establish and improve incident response, escalation, troubleshooting, and post-incident review processes.
  • Partner with Software Engineering, Product, Cloud/Platform, Security, and Infrastructure teams.
  • Apply software engineering principles to automate and solve reliability and operational challenges.
  • Improve system availability, scalability, resilience, performance, and operational efficiency.
  • Drive CI/CD improvements and deployment reliability.
  • Architect and support applications and infrastructure running on high-growth cloud platforms.
  • Lead reliability programs across multiple B2B SaaS products.
  • Establish engineering metrics and use data/KPIs to identify and resolve operational gaps.
  • Drive automation and reduction of repetitive operational work.
  • Mentor engineering leaders and senior SRE engineers.
  • Communicate reliability strategy, risks, metrics, and business impact to senior leadership.
  • Influence engineering teams and cross-functional stakeholders to adopt reliability best practices.

Bachelor's degree in Computer Science, Information Systems, Engineering, or a related field.

  • 18+ years of experience in Software Engineering, SRE, Site Reliability, Platform Engineering, or related reliability/engineering roles.
  • 5+ years of leadership experience at Director level or equivalent.
  • Strong experience in SaaS / B2B SaaS / cloud product companies.
  • Proven ability to apply software engineering principles and practices to solve reliability and operational challenges.
  • Strong expertise in SLI/SLO, monitoring, observability, and reliability engineering.
  • Strong experience with CI/CD and modern software delivery practices.
  • Strong experience with incident response, problem management, RCA, and production operations.
  • Strong AWS expertise.
  • Experience with container orchestration, such as Kubernetes.
  • Experience leading reliability programs across multiple SaaS products.
  • Experience architecting applications or infrastructure for high-growth cloud platforms.
  • Experience with large-scale distributed systems in B2B SaaS environments.
  • Strong leadership, communication, stakeholder management, and influencing skills.
  • Demonstrated experience driving operational excellence through metrics, KPIs, SLOs, and reliability objectives.
Preferred Skills
  • Kubernetes / containerized environments
  • AWS cloud architecture
  • Infrastructure and application observability
  • Distributed systems
  • Microservices
  • Infrastructure automation
  • Infrastructure-as-Code
  • Terraform
  • Prometheus / Grafana
  • Datadog / New Relic / Splunk or similar observability platforms
  • CI/CD platforms
  • Disaster recovery and business continuity
  • Capacity and performance engineering
  • Chaos/resilience engineering

Skills: kubernetes,large-scale distributed systems,metrics,container orchestration,production operations,kpis,cloud,multiple saas products,problem management,sli/slo,leadership,ci/cd,b2b saas environments,observability,slos,aws,saas,b2b,monitoring,reliability engineering

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Director – Site Reliability Engineering (SRE)
Director – Site Reliability Engineering (SRE)

Umanist NA • Navi Mumbai

On-site
INR 6,000,000 - 12,000,000
Director – Site Reliability Engineering (SRE)
Director – Site Reliability Engineering (SRE)

Umanist NA • Mumbai

On-site
INR 6,000,000 - 9,000,000
Site Reliability Engineering (SRE) Manager
Site Reliability Engineering (SRE) Manager

Acesoft Labs • Hyderabad

Hybrid
INR 4,000,000 - 6,000,000
SRE Monitoring & Observability
SRE Monitoring & Observability

Advance Career Solutions • Chennai District

Hybrid
INR 1,200,000 - 2,000,000
Software Engineer
Software Engineer

PwC • Hyderabad, Bengaluru

Hybrid
INR 2,800,000 - 5,200,000
Senior Consultant - Site Reliability Engineer
Senior Consultant - Site Reliability Engineer

HCA Healthcare • Hyderabad

On-site
INR 3,000,000 - 5,200,000
SRE / Production Engineering
SRE / Production Engineering

Infosys • Bengaluru

On-site
INR 3,500,000 - 7,000,000
Site Reliability Engineer
Site Reliability Engineer

Mumba Technologies, Inc. • Gurugram District

Hybrid
INR 1,500,000 - 2,100,000
Site Reliability Engineer
Site Reliability Engineer

PwC Acceleration Center India • Bengaluru

On-site
INR 2,200,000 - 3,800,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

TymblHub • India

On-site
INR 2,500,000 - 4,000,000