Site Reliability Engineer Lead

Cvent

Gurugram District

Hybrid

INR 3,500,000 - 6,000,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Cvent is seeking a Lead Site Reliability Engineer to scale our cloud platform, improve stability, and accelerate deployments. You will guide reliability strategy, implement scalable AWS architectures, and drive CI/CD automation while mentoring engineers.

You will shape AI-augmented operations, advance DevOps practices, and partner with engineering to ensure secure, compliant, and efficient systems. This role requires strong leadership and hands-on expertise in distributed systems.

Qualifications

  • 8–11 years in Site Reliability, DevOps, or Cloud Engineering with leadership experience.
  • Deep AWS expertise across multi-region architectures and networking.
  • Advanced development/scripting (Python, Go, Bash) for automation.
  • Mastery of CI/CD platforms and containerization (Docker, Kubernetes).
  • Experience with SLIs/SLOs, error budgets, and incident postmortems.

Responsibilities

  • Set direction and long-term strategy for solving complex reliability problems.
  • Lead design and implementation of large-scale AWS architectures for reliability and scale.
  • Develop and refine CI/CD pipelines and automated deployments.
  • Drive error budgets and define organization-wide reliability targets (SLIs/SLOs).
  • Lead deep-dive RCAs for major incidents using Datadog, Prometheus, Grafana, ELK; conduct blameless postmortems.

Skills

SRE principles
DevOps culture
AWS expertise
Python
Go
Bash
CI/CD
Docker
Kubernetes
Terraform
CloudFormation
AWS CDK
Jenkins
GitHub Actions
Argo CD
Datadog
Prometheus
Grafana
ELK
Unix/Linux
Incident management

Tools

Datadog
Prometheus
Grafana
ELK

Job description

Lead Site Reliability Engineer
Company Overview

Cvent, Inc. (www.cvent.com) is the worlds leading provider of cloud-based software for meetings and event management. Our platform of products includes software to manage and facilitate online event registration, meeting site selection, event management, e‑mail marketing and web surveys. We also develop mobile apps for both corporate and consumer events. Founded in 1999, we currently have 5000+ talented and dedicated employees and are headquartered just outside of Washington, D.C., in McLean, Virginia. Cvent has received several awards and honors recognizing our strong company culture, innovative products, stellar customer service and support, visionary leadership and investment in our employees. We currently have job openings across all departments and locations and are looking to add valuable team members to further strengthen the companys DNA.

Job Description:

Cvent is looking for a Lead Site Reliability Engineer to help us scale our systems and ensure stability, reliability and performance and rapid deployments of our platform. We build teams that are inclusive, collaborative, and have a strong sense of ownership for the things they build. If you have a passion and track record for solving problems; moreover, have strong leadership skills, this is a great fit for you.

As a Lead Engineer, you will demonstrate both emerging and current technologies, methods, and processes contributing to the evolution of software deployment processes, enhancing security, reducing risk, and improving the overall end‑user experience. As part of the Technology R&D Team, you will play an integral part in advancing DevOps maturity and be a part of a new culture of quality and site reliability. You will continually improve reliability, resiliency and scalability of our products, processes, and procedures. In this position, you would also be expected to ramp up to manage/mentor engineers and ensure their technical growth.

AI at Cvent: Leading the Future:

Are you ready to shape the future of work at the intersection of human expertise and AI innovation? At Cvent, we’re committed to continuous learning and adaptation—AI isn’t just a tool for us, it’s part of our DNA. We’re looking for candidates who are eager to evolve alongside technology. If you love to experiment boldly, share your discoveries, and

help define best practices for AI-augmented work, you’ll thrive here. Our team values professionals who thoughtfully integrate AI into their daily work, delivering exceptional results while relying on the human judgment and creativity that drive real innovation. Throughout our interview process, you’ll have the chance to demonstrate how you use AI to learn, iterate, and amplify your impact. If you’re excited to be part of a team that’s leading the way in AI-powered collaboration, we’d love to meet you.

What You Will Be Doing
  • Set the direction and long-term strategy for solving complex problems; communicate timeline, scope, risks, and the technical roadmap to leadership and stakeholders.
  • Keep abreast of emerging cloud technologies, running POCs to assess suitability and value.
  • Lead design and implementation of large-scale AWS architectures optimized for reliability, scalability, and cost.
  • Develop and refine CI/CD pipelines and automated deployments for complex environments.
  • Drive error budgets; define and track SLIs/SLOs; own organization-wide reliability targets.
  • Lead deep-dive RCAs for major incidents using Datadog, Prometheus, Grafana, and ELK; run blameless postmortems.
  • Establish best practices for containerization (Docker, Kubernetes) and infrastructure as code (Terraform, AWS CDK)
  • Mentor, coach, and provide technical escalation for SRE/DevOps engineers; foster a learning and ownership culture.
  • Direct AI/automation initiatives to improve deployment, monitoring, troubleshooting, and developer efficiency (e.g., ChatGPT, Copilot, generative AI for runbooks and incident management)
  • Partner with engineering, product, and business to shape reliability strategy, incident process, architecture reviews, and roadmap.
  • Ensure security, regulatory, and operational compliance across cloud architecture and automation.
  • continuously communicating timeline, scope, risks, and technical road map.
What You Need for this Position
  • 8-11 years in Site Reliability, DevOps, or Cloud Engineering, including substantial leadership experience.
  • Strong understanding of SRE principles and DevOps culture; adept at applying software engineering tools, methods, and practices
  • Deep AWS expertise across advanced networking, identity (IAM), and multi-region architectures; experience with large-scale, multi-tier distributed systems.
  • Advanced development and scripting skills (Python, Go, Bash) focused on automation and troubleshooting distributed systems.
  • Mastery of CI/CD platforms (Jenkins, GitHub Actions, Argo CD), containerization (Docker, Kubernetes), and infrastructure as code (Terraform, CloudFormation, AWS CDK).
  • Proven track record driving SLIs/SLOs, error budgets, and reliability targets at scale, plus leading complex incident processes and blameless postmortems.
  • Strong Unix/Linux background with deep knowledge of system internals.
  • Experience with database technologies (SQL Server, PostgreSQL, Couchbase preferred).
  • Effective people leader and mentor with excellent communication and stakeholder management skills; bias for execution.
  • Direct AI/automation initiatives to improve deployment, monitoring, troubleshooting, and developer efficiency (e.g., ChatGPT, Copilot, generative AI for runbooks and incident management).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead SRE
Lead SRE

Cvent • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Namely • India

On-site
INR 1,500,000 - 2,500,000
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
Lead SRE
Lead SRE

Cvent, Inc. • India

On-site
INR 2,500,000 - 4,500,000
Lead Site Reliability Engineer (Security)
Lead Site Reliability Engineer (Security)

Cvent • Gurugram District

On-site
INR 4,000,000 - 6,000,000
Lead Site Reliability Engineer (Security)
Lead Site Reliability Engineer (Security)

Cvent, Inc. • Gurugram District

Hybrid
INR 1,800,000 - 2,400,000
Lead Site Reliability Engineer (Security)
Lead Site Reliability Engineer (Security)

Namely • India

On-site
INR 1,400,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Cvent • Bengaluru

On-site
INR 3,500,000 - 6,500,000
Lead Security Engineer
Lead Security Engineer

Cvent • Gurugram District, Bengaluru

Hybrid
INR 3,000,000 - 6,000,000
Site Reliability Engineer Lead (Immediate Joiner)
Site Reliability Engineer Lead (Immediate Joiner)

HiLabs • Pune District

On-site
INR 2,500,000 - 4,200,000