Director, Site Reliability Engineering

Vertafore Career Center

Denver (CO)

On-site

USD 175,000 - 220,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Vertafore seeks a Director of Site Reliability Engineering to lead reliability, performance, and observability across a portfolio of products. You will define SLIs/SLOs, direct release engineering, and drive automation and CI/CD practices with teams of managers and engineers.

You will manage on-call rotations, balance error budgets, and collaborate with Cloud Ops, Security, and Product Development to meet SLAs and scale the platform.

Qualifications

  • Experience leading SRE teams and driving reliability initiatives.
  • Proven ability to design CI/CD pipelines and automate repetitive work.
  • On-call leadership with blameless incident postmortems.
  • Ability to align reliability goals with product roadmaps.

Responsibilities

  • Define and enforce SLIs/SLOs for key Vertafore products.
  • Oversee CI/CD pipelines and drive toil reduction through automation.
  • Manage on-call rotations and incident response across portfolios.
  • Collaborate with Cloud Ops, InfoSec, and product teams to meet SLAs.

Skills

Leadership
Incident management
Automation & observability
Cross-functional collaboration
Team mentoring
Capacity planning

Tools

GitLab
Jenkins
Ansible
LaunchDarkly

Job description

$175,000 - $220,000 / year + Bonus

The insurance industry runs on Vertafore. We equip agencies, MGAs, and carriers with the core digital systems, specialized AI, and data-driven foundation to eliminate distribution drag across the insurance lifecycle, spanning sales, servicing, and back-office operations.

Underpinned by unmatched speed and performance power, we are the trusted backbone that's taking the insurance industry from friction to flow with Distribution Velocity - speed, performance, and trust - to drive growth at scale.

With over 95% of the top agencies and insurers and 50% of industry compliance transactions running through Vertafore, we lead at the intersection of innovation and trust, giving insurance professionals the confidence to transform and win in the AI era.

Our reach is global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India.

The Director, Site Reliability Engineering (SRE) will lead reliability, performance, and observability initiatives for a portfolio of Vertafore products. This role owns SLIs/SLOs, incident response, automation, and CI/CD practices for assigned product families. Directors will manage multiple teams and collaborate with Product Development, Architecture, Cloud Operations, Information Security, and other SRE leaders to ensure operational excellence. This role is responsible for bridging the gap between development and operations by applying a software engineering mindset to system administration. You will own the lifecycle of services - from inception and design, through deployment, operation, and refinement.

Key Responsibilities
  • Product Reliability Leadership Define and enforce SLIs/SLOs for a subset of Vertafore flagship products. Drive observability strategy across application and infrastructure layers.
  • Release Engineering & Toil Reduction Oversee CI/CD pipelines for product deployments using tools like GitLab, Jenkins, Ansible, LaunchDarkly. Monitor and cap "Toil" (manual, repetitive operational work) at 50% using Automation and AI tools, ensuring the team spends the remaining time on project work that scales the system.
  • Error Budget Management Manage "Error Budgets" to balance the velocity of feature releases with the stability of the platform, ensuring clear consequences when budgets are exhausted.
  • Incident Management Define and participate in 24x7 on-call rotations for assigned products; ensure rapid resolution and blameless postmortems.
  • Cross-Functional Collaboration Partner with Cloud Ops on capacity planning, OS patching (app tier), and load balancing (ALB, F5). Align reliability goals with product roadmaps and customer SLAs.
  • Team Leadership Manage a group of Managers and Engineers, mentor teams on automation, observability, and reliability best practices.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Vertafore Career Center • Denver (CO)

On-site
USD 110,000 - 145,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Vertafore Career Center • Denver (CO)

On-site
USD 160,000 - 180,000
R Site Reliability Engineer Ii Restaurant365 Denver, Colorado, US
R Site Reliability Engineer Ii Restaurant365 Denver, Colorado, US

Artha Nexgen • Town of Texas (WI), Northern (KY)

Hybrid
USD 180,000 - 260,000
Director, Cloud Operations
Director, Cloud Operations

Hirebridge • Denver (CO)

On-site
USD 190,000 - 220,000
T Sr. Site Reliability Engineer Tiger Analytics Washington, Dc, US
T Sr. Site Reliability Engineer Tiger Analytics Washington, Dc, US

Artha Nexgen • Washington, Northern (KY)

Hybrid
USD 140,000 - 190,000
Director, Cloud Operations
Director, Cloud Operations

Vertafore Career Center • Colorado

On-site
USD 190,000 - 220,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Vertafore • Denver (CO)

On-site
USD 110,000 - 145,000
Medical, vision & dental plans
401(k) Retirement Savings Plan with
Life, LTD/AD&D
+3
Director, Reliability & Observability Engineering
Director, Reliability & Observability Engineering

Vertafore Career Center • Denver (CO)

On-site
USD 175,000 - 220,000
Sr. Software Engineer
Sr. Software Engineer

Vertafore Career Center • Windsor (CT)

On-site
USD 90,000 - 125,000
Director, Development
Director, Development

Vertafore Career Center • Denver (CO)

On-site
USD 155,000 - 200,000