Lead SRE

Jobgether

India

On-site

INR 3,000,000 - 6,500,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Mentorship programs
Structured leadership development
Generous paid time off
Global exposure and cross-team impact

Job summary

Jobgether is seeking a Lead Site Reliability Engineer based in India to lead a high-performing SRE team across multiple time zones. You will drive reliability, scalability and performance of cloud-native platforms while embedding SRE best practices and data-driven operations.

You will collaborate with engineering, product and infrastructure teams to reduce toil, improve observability and implement CI/CD, IaC, and incident response processes.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or related field or equivalent experience.
  • 10+ years in software engineering/DevOps/SRE with at least 3 years in leadership.
  • Experience managing large-scale distributed systems in cloud-native environments (AWS or Azure).
  • Strong expertise in monitoring/observability (AppDynamics, Splunk).
  • Experience with Terraform, Ansible; CI/CD with Jenkins and GitHub Actions.
  • Deep knowledge of Linux, networking, containers (Docker, Kubernetes).
  • Excellent communication, collaboration and stakeholder-management skills.

Responsibilities

  • Lead and mentor a global SRE team across multiple time zones.
  • Execute the SRE roadmap aligned with business objectives.
  • Improve system reliability, scalability, performance and resilience.
  • Drive adoption of SRE principles, including SLIs, SLOs and error budgets.
  • Lead capacity planning, incident response, RCAs and blameless postmortems.
  • Champion automation and infrastructure as code to reduce toil.
  • Build and maintain CI/CD pipelines and observability capabilities.

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead SRE based in India.

This role offers the opportunity to lead and scale a high-performing Site Reliability Engineering function supporting critical production systems. You will be responsible for strengthening reliability, scalability, performance, and operational resilience across cloud-native platforms. The position combines technical leadership with strategic planning, team development, automation, observability, and incident management. You will guide SRE teams across multiple time zones while partnering closely with engineering, product and infrastructure stakeholders. A major focus will be on embedding SRE best practices, reducing operational toil, and driving continuous improvement. You will help establish a culture of operational excellence, learning and reliable service delivery across a global technology environment.

Accountabilities:

  • Lead, mentor and develop a high-performing team of SREs working across multiple time zones, fostering strong collaboration, technical excellence, and continuous learning.
  • Execute the SRE roadmap in alignment with broader business and engineering objectives, ensuring reliability priorities are clearly defined and delivered.
  • Partner with engineering, product, and infrastructure teams to improve system reliability, scalability, performance and operational resilience.
  • Drive adoption of SRE principles, including Service Level Indicators, Service Level Objectives, error budgets, reliability metrics, and data-driven operational practices.
  • Lead capacity planning, performance optimization, disaster recovery, incident response, root cause analysis and blameless postmortems for critical systems.
  • Champion automation and infrastructure improvements to reduce operational toil and increase engineering efficiency.
  • Build and maintain CI/CD pipelines, observability capabilities, infrastructure-as-code solutions and operational tooling while ensuring alignment with security, privacy and regulatory requirements.
  • Establish and monitor key metrics covering system health, service reliability, operational performance and team effectiveness.
Requirements:
  • Hold a bachelor’s degree in Computer Science, Engineering, or a related field, or demonstrate equivalent practical experience.
  • Bring 10+ years of experience across software engineering, DevOps, or Site Reliability Engineering, including at least 3 years in a technical leadership or people-management capacity.
  • Demonstrate proven experience managing large-scale, distributed systems within cloud-native environments such as AWS or Azure.
  • Possess strong expertise in monitoring and observability technologies such as AppDynamics and Splunk, as well as automation tools including Terraform and Ansible.
  • Have strong experience designing and managing CI/CD pipelines using technologies such as Jenkins and GitHub Actions.
  • Demonstrate deep knowledge of Linux systems, networking, containers, Docker, Kubernetes and modern infrastructure engineering practices.
  • Possess excellent communication, collaboration, leadership and stakeholder-management skills, with the ability to influence teams across a complex organization.
  • Experience working in a global, matrixed organization and contributions to open-source SRE or DevOps tools are considered valuable.
Benefits:
  • Opportunities to lead and shape a growing SRE function within a global, technology-driven environment.
  • Professional development opportunities, including structured leadership and manager-development programs.
  • Access to mentorship opportunities supporting technical growth, leadership development and career progression.
  • Generous paid time off and Volunteer Time Off opportunities, subject to applicable eligibility requirements.
  • Competitive healthcare and employee benefits, including medical and dental coverage.
  • Access to a Global Employee Assistance Program providing additional support and resources.
  • Opportunities to participate in Employee Impact Groups and employee experience initiatives that encourage connection, inclusion, collaboration and community involvement.
  • The opportunity to work with distributed teams across multiple time zones while contributing to critical reliability and infrastructure initiatives.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
Lead SRE
Lead SRE

Cvent, Inc. • India

On-site
INR 2,500,000 - 4,500,000
Senior Manager - Site Reliability Engineer|NR-2026-0246
Senior Manager - Site Reliability Engineer|NR-2026-0246

Media.net • Bengaluru

On-site
INR 6,000,000 - 8,000,000
SRE Practice Lead
SRE Practice Lead

Coforge • Dadri

On-site
INR 3,000,000 - 4,200,000
Senior SRE
Senior SRE

CloudRaft • India

On-site
INR 2,500,000 - 4,500,000
Competitive salary
Premium health insurance & wellness
AI stack & GPU infrastructure
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities
SRE Lead
SRE Lead

Manatal • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Senior SRE Engineer
Senior SRE Engineer

EPAM Systems • Gurugram District

On-site
INR 3,000,000 - 5,000,000
Lead SRE
Lead SRE

Cvent • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Hilabs • Pune District

On-site
INR 1,500,000 - 2,500,000