Senior Site Reliability Engineer

Cvent, Inc.

Tysons (VA)

Hybrid

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cvent, Inc. is looking for a Senior Site Reliability Engineer in Tysons, Virginia. The role involves using advanced development and operations skills to enhance infrastructure and guide development teams in producing reliable systems.

The ideal candidate will have strong scripting abilities, experience with cloud services like AWS, and a commitment to leveraging AI for improved operations. This hybrid position includes 2 days in the office and is integral to driving continuous improvement in Cvent’s technology processes.

Qualifications

  • Experience with Agile SDLC methodologies and championing CI/CD.
  • Scripting languages like Ruby, Groovy, Bash, PowerShell, Typescript, or Python.
  • Exposure to managing AWS services / operational knowledge of managing applications in AWS.
  • Experience with configuration management tools such as Chef, Puppet, Ansible or equivalent.
  • Hands-on experience with Windows and Linux/Unix administration.
  • Working with APM, monitoring, and logging tools (Datadog, New Relic, Splunk).
  • Good understanding of containerization concepts – Docker, ECS, EKS, Kubernetes.
  • Hands-on experience with AI coding assistants such as Claude Code, and AI agents for operational and engineering workflows.

Responsibilities

  • Enlighten, enable and empower multi-disciplinary teams across multiple applications and locations.
  • Guide development teams through infrastructure decisions.
  • Run Infrastructure as Code applications.
  • Implement and maintain SLIs/SLOs and conduct incident retrospectives.
  • Focus on automation of tasks and AI-driven operational efficiency.

Skills

Experience with Agile SDLC methodologies
Scripting languages (Ruby, Groovy, Bash, PowerShell, Typescript, Python)
Managing AWS services
Configuration management tools (Chef, Puppet, Ansible)
Windows and Linux/Unix administration
APM, monitoring, and logging tools (Datadog, New Relic, Splunk)
Containerization concepts (Docker, ECS, EKS, Kubernetes)
Build tools (Jenkins)
AI coding assistants (Claude Code)
AI/ML applications for observability

Job description

Overview
Our Culture and Impact

Cvent is a leading meetings, events, and hospitality technology provider with more than 5,500+ employees and ~30,000 customers worldwide, including 60% of the Fortune 500. Founded in 1999, Cvent delivers a comprehensive event marketing and management platform for marketers and event professionals and offers software solutions to hotels, special event venues and destinations to help them grow their group/MICE and corporate travel business. Our technology brings millions of people together at events around the world. In short, we’re transforming the meetings and events industry through innovative technology that powers the human connection.

Cvent's strength lies in its people, fostering a culture where everyone is encouraged to think like entrepreneurs, taking risks and making decisions confidently. We value diverse perspectives and celebrate differences, working together with colleagues and clients to build strong connections.

AI at Cvent: Leading the Future

Are you ready to shape the future of work at the intersection of human expertise and AI innovation? At Cvent, we’re committed to continuous learning and adaptation—AI isn’t just a tool for us, it’s part of our DNA. We’re looking for candidates who are eager to evolve alongside technology. If you love to experiment boldly, share your discoveries, and help define best practices for AI-augmented work, you’ll thrive here. Our team values professionals who thoughtfully integrate AI into their daily work, delivering exceptional results while relying on the human judgment and creativity that drive real innovation.

Throughout our interview process, you’ll have the chance to demonstrate how you use AI to learn, iterate, and amplify your impact. If you’re excited to be part of a team that’s leading the way in AI-powered collaboration, we’d love to meet you.

As a Senior Site Reliability Engineer, you will use your advanced development and operations knowledge to run Infrastructure as Code applications, build pipelines, and enable development teams. You will guide development teams through infrastructure decisions, conduct incident retrospectives, and implement and maintain SLIs/SLOs. You will help teams evaluate their reliability posture and prioritize work to solve reliability issues, identify and prioritize issues, find universal solutions to common problems, and mentor and support junior staff. You will put AI at the center of how we work – using AI tools and agents to automate toil, accelerate incident response, and build smarter, self‑healing systems. As a Cvent SRE you will be a force for positive change and drive continuous improvement.

In This Role, You Will:
  • Enlighten, enable and empower a fast‑growing set of multi‑disciplinary teams, across multiple applications and locations.
  • Guide development teams through infrastructure decisions and help them evaluate their reliability posture and prioritize work to solve reliability issues.
  • Run Infrastructure as Code applications, build pipelines, and enable development teams.
  • Tackle complex development, automation and business process problems.
  • Champion Cvent standards and best practices.
  • Ensure the scalability, performance, and resilience of our suite of products.
  • Work with the development and product team of a new application to establish the right monitoring and alerting strategy.
  • Implement and maintain SLIs/SLOs and conduct incident retrospectives.
  • Develop build, test and deployment automation that seamlessly targets multiple on‑premises and AWS regions.
  • Help a dev team working on a legacy code base to realize zero‑down‑time deployments.
  • Focus on automation of tasks and AI‑driven operational efficiency – automate all the things!
  • Leverage AI tools and agents (e.g., Claude) to accelerate development, reduce toil, and speed up incident response.
  • Build and integrate AI‑driven automation into monitoring, alerting, and remediation workflows – such as anomaly detection, intelligent runbooks, and auto‑remediation.
  • Apply AI/ML techniques to observability, log analysis, and capacity planning to surface issues earlier and resolve them faster.
  • Evaluate, pilot, and champion emerging AI tooling, sharing best practices and helping development teams adopt AI responsibly and effectively.
  • Give back by working on and contributing to Open‑Source projects.
Here's What You Need:
  • Experience with Agile SDLC methodologies and championing CI/CD.
  • Scripting languages like Ruby, Groovy, Bash, PowerShell, Typescript, or Python.
  • Exposure to managing AWS services / operational knowledge of managing applications in AWS.
  • Experience with configuration management tools such as Chef, Puppet, Ansible or equivalent.
  • Hands‑on experience with Windows and Linux/Unix administration.
  • Working with APM, monitoring, and logging tools (Datadog, New Relic, Splunk).
  • Good understanding of containerization concepts – Docker, ECS, EKS, Kubernetes.
  • Experience managing 3‑tier application stacks.
  • Experience with build tools such as Jenkins.
  • Working experience with NoSQL databases such as MongoDB, Couchbase, Postgres, etc.
  • F5 load balancing concepts.
  • Understanding of basic networking concepts.
  • Experience with package managers such as Nexus, Artifactory or equivalent.
  • Hands‑on experience with AI coding assistants such as Claude Code, and AI agents for operational and engineering workflows.
  • Familiarity with applying AI/ML to observability, anomaly detection, log analysis, or automated remediation.
  • Experience integrating LLM‑based or AI‑driven tools into engineering and CI/CD pipelines (preferred).
  • A growth mindset, with a commitment to staying current with AI advancements through self‑driven learning.
  • Strong observability practices.
  • Understanding of SRE concepts and the DevOps culture, with a focus on leveraging software engineering tools, methodologies, and concepts.
  • Intermediate troubleshooting skills and the ability to lead incident response efforts.
  • Identifying and prioritizing issues and finding solutions to common problems through a holistic view of Cvent systems.
  • Mentoring and supporting junior staff through strong communication skills.

Hybrid: 2 days in office

We are not able to offer sponsorship for this position

Physical Demands

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Namely • United States

Hybrid
USD 120,000 - 150,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Cvent • Tysons (VA)

Hybrid
USD 110,000 - 140,000
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

R&D • United States

On-site
USD 90,000 - 140,000
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

Calabrio • United States

On-site
USD 90,000 - 130,000
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

Verint Systems, Inc. • Columbus (OH)

On-site
USD 85,000 - 110,000
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

Verint Systems, Inc. • Richmond (VA)

On-site
USD 90,000 - 120,000
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

Verint Systems, Inc. • Annapolis (MD)

On-site
USD 100,000 - 140,000
Application Support Engineer
Application Support Engineer

Cvent • McLean (VA)

On-site
USD 85,000 - 120,000
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

Verint Systems, Inc. • Frankfort (KY)

Hybrid
USD 65,000 - 90,000
Head of SRE
Head of SRE

Wand AI • Palo Alto (CA)

On-site
USD 130,000 - 180,000