Lead SRE

Cvent

Gurugram District

On-site

INR 4,000,000 - 7,000,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cvent, Inc. is seeking a Lead Site Reliability Engineer to set the reliability strategy, own large-scale AWS architectures, and lead CI/CD automation. You will mentor SRE/DevOps engineers and partner with product and engineering to improve deployment, monitoring, and incident response.

Ideal candidates have 7–10 years in SRE/DevOps, strong AWS multi-region expertise, and deep scripting skills in Python/Go. This is an on-site role in Gurugram, India, requiring leadership and broad collaboration.

Qualifications

  • 7-10 years in Site Reliability, DevOps, or Cloud Engineering, including leadership experience.
  • Strong AWS expertise across multi-region architectures and advanced networking.
  • Deep scripting and automation focus (Python, Go, Bash) for distributed systems.

Responsibilities

  • Set the direction and long-term strategy for solving complex problems; communicate timeline, scope, risks, and the technical roadmap to leadership and stakeholders.
  • Lead design and implementation of large-scale AWS architectures optimized for reliability, scalability, and cost.
  • Develop and refine CI/CD pipelines and automated deployments for complex environments.
  • Drive error budgets; define and track SLIs/SLOs; own organization-wide reliability targets.
  • Lead deep-dive RCAs for major incidents using Datadog, Prometheus, Grafana, and ELK; run blameless postmortems.
  • Establish best practices for containerization (Docker, Kubernetes) and infrastructure as code (Terraform, AWS CDK).
  • Mentor, coach, and provide technical escalation for SRE/DevOps engineers; foster a learning and ownership culture.
  • Direct AI/automation initiatives to improve deployment, monitoring, troubleshooting, and developer efficiency.

Skills

SRE
DevOps
Cloud engineering
Leadership
Automation
Incident management
Python
Go
Bash

Tools

AWS
Docker
Kubernetes
Terraform
CloudFormation
AWS CDK
Jenkins
GitHub Actions
Argo CD
Datadog
Prometheus
Grafana
ELK

Job description

Company Overview
Cvent, Inc. (www.cvent.com) is the world’s leading provider of cloud-based software for meetings and event management. Our platform of products includes software to manage and facilitate online event registration, meeting site selection, event management, e-mail marketing and web surveys. We also develop mobile apps for both corporate and consumer events. Founded in 1999, we currently have 5000+ talented and dedicated employees and are headquartered just outside of Washington, D.C., in McLean, Virginia. Cvent has received several awards and honors recognizing our strong company culture, innovative products, stellar customer service and support, visionary leadership and investment in our employees. We currently have job openings across all departments and locations and are looking to add valuable team members to further strengthen the company’s DNA.

Lead Site Reliability Engineer
Company Overview

Cvent, Inc. (www.cvent.com) is the world’s leading provider of cloud-based software for meetings and event management. Our platform of products includes software to manage and facilitate online event registration, meeting site selection, event management, e-mail marketing and web surveys. We also develop mobile apps for both corporate and consumer events. Founded in 1999, we currently have 5000+ talented and dedicated employees and are headquartered just outside of Washington, D.C., in McLean, Virginia. Cvent has received several awards and honors recognizing our strong company culture, innovative products, stellar customer service and support, visionary leadership and investment in our employees. We currently have job openings across all departments and locations and are looking to add valuable team members to further strengthen the company’s DNA.

Set the direction and long-term strategy for solving complex problems; communicate timeline, scope, risks, and the technical roadmap to leadership and stakeholders.

Keep abreast of emerging cloud technologies, running POCs to assess suitability and value.

Lead design and implementation of large-scale AWS architectures optimized for reliability, scalability, and cost.

Develop and refine CI/CD pipelines and automated deployments for complex environments.

Drive error budgets; define and track SLIs/SLOs; own organization-wide reliability targets.

Lead deep-dive RCAs for major incidents using Datadog, Prometheus, Grafana, and ELK; run blameless postmortems.

Establish best practices for containerization (Docker, Kubernetes) and infrastructure as code (Terraform, AWS CDK).

Mentor, coach, and provide technical escalation for SRE/DevOps engineers; foster a learning and ownership culture.

Direct AI/automation initiatives to improve deployment, monitoring, troubleshooting, and developer efficiency (e.g., ChatGPT, Copilot, generative AI for runbooks and incident management).

Partner with engineering, product, and business to shape reliability strategy, incident process, architecture reviews, and roadmap.

Ensure security, regulatory, and operational compliance across cloud architecture and automation.

continuously communicating timeline, scope, risks, and technical road map.

  • 7-10 years in Site Reliability, DevOps, or Cloud Engineering, including substantial leadership experience.

Strong understanding of SRE principles and DevOps culture; adept at applying software engineering tools, methods, and practices.

Deep AWS expertise across advanced networking, identity (IAM), and multi-region architectures; experience with large-scale, multi-tier distributed systems.

Advanced development and scripting skills (Python, Go, Bash) focused on automation and troubleshooting distributed systems.

Mastery of CI/CD platforms (Jenkins, GitHub Actions, Argo CD), containerization (Docker, Kubernetes), and infrastructure as code (Terraform, CloudFormation, AWS CDK).

Proven track record driving SLIs/SLOs, error budgets, and reliability targets at scale, plus leading complex incident processes and blameless postmortems.

Strong Unix/Linux background with deep knowledge of system internals.

Experience With Database Technologies (SQL Server, PostgreSQL, Couchbase Preferred).

Effective people leader and mentor with excellent communication and stakeholder management skills; bias for execution.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead SRE
Lead SRE

Cvent, Inc. • India

On-site
INR 2,500,000 - 4,500,000
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Namely • India

On-site
INR 1,500,000 - 2,500,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Technologies Pvt. Ltd. • Pune District

On-site
INR 900,000 - 1,400,000
Lead Site Reliability Engineer (Security)
Lead Site Reliability Engineer (Security)

Cvent • Gurugram District

On-site
INR 4,000,000 - 6,000,000
Lead Site Reliability Engineer (Security)
Lead Site Reliability Engineer (Security)

Namely • India

On-site
INR 1,400,000 - 2,000,000
Lead Site Reliability Engineer (Security)
Lead Site Reliability Engineer (Security)

Cvent, Inc. • Gurugram District

Hybrid
INR 1,800,000 - 2,400,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
SRE Lead
SRE Lead

Manatal • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer Lead (Immediate Joiner)
Site Reliability Engineer Lead (Immediate Joiner)

HiLabs • Pune District

On-site
INR 2,500,000 - 4,200,000