Sr. Site Reliability Engineer

Vertafore Career Center

Denver (CO)

On-site

USD 110,000 - 145,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Vertafore is seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services in Denver, CO. You will oversee the full-service lifecycle—from design and deployment readiness through production operations and incident response.

This role emphasizes strong software engineering, autonomous operation across AWS, hybrid data centers, and customer-hosted environments, with a focus on observability, fault tolerance,

Qualifications

  • 8+ years of hands-on Site Reliability Engineering or reliability-focused engineering experience with end-to-end service ownership.
  • Senior engineering scope with accountability for reliability outcomes.
  • Strong software engineering skills in C#, .NET, Java, Python, React, or similar technologies.
  • Practical experience applying SRE principles (SLIs, SLOs, error budgets).
  • Hands-on experience with AWS, Kubernetes, CI/CD, infrastructure as code and hybrid environments.
  • Strong knowledge of Linux and Windows systems, application platforms and relational databases.
  • Bachelor's or master's degree in computer science or equivalent experience.
  • Participation in anonymous-call rotation; flexible hours as required.

Responsibilities

  • Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health.
  • Define and manage SLIs and SLOs, using error budgets to guide delivery decisions.
  • Influence of service and system design to improve fault tolerance, observability and operational sustainability.
  • Debug complex production issues across application code, services and infrastructure using software engineering practices.
  • Perform root cause analysis using logs, metrics, traces, and code-level investigation.
  • Build automation and self-healing mechanisms to prevent repeat failures.
  • Execute production changes (patching, certificate management, software releases) with safety, automation, and observability.
  • Design and operate production observability aligned to service health and customer impact.
  • Lead and participate in incident response, for high-severity events.
  • Collaborate with engineering, product, architecture, and operations teams.
  • Operate with autonomy and sound judgment in reliability decisions.

Skills

C#
Java
Python
.NET
React
SRE principles
Observability
Incident response
AWS
Kubernetes
CI/CD
Linux/Windows

Education

Bachelor's or Master's in CS or equivalent

Tools

Terraform
CI/CD pipelines
AWS Console
Git

Job description

$110,000 - $145,000 / year + Bonus

The insurance industry runs on Vertafore. We equip agencies, MGAs, and carriers with the core digital systems, specialized AI, and data-driven foundation to eliminate distribution drag across the insurance lifecycle, spanning sales, servicing, and back-office operations.

Underpinned by unmatched speed and performance power, we are the trusted backbone that is taking the insurance industry from friction to flow with Distribution Velocity - speed, performance, and trust - to drive growth at scale.

With over 95% of the top agencies and insurers and 50% of industry compliance transactions running through Vertafore, we lead at the intersection of innovation and trust, giving insurance professionals the confidence to transform and win in the AI era.

Our reach is global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India.

We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is accountable for the full-service lifecycle, from design and deployment readiness through production operations, incident response, and continuous improvement. Reliability is a core engineering responsibility, requiring strong software engineering skills and autonomous operation across AWS, hybrid data centers, and customer-hosted environments.

Roles and Responsibilities
  • Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health.
  • Define and manage SLIs and SLOs, using error budgets to guide delivery decisions.
  • Influence of service and system design to improve fault tolerance, observability and operational sustainability.
  • Debug complex production issues across application code, services and infrastructure using software engineering practices.
  • Perform root cause analysis using logs, metrics, traces, and code-level investigation.
  • Build automation and self-healing mechanisms to prevent repeat failures.
  • Execute production changes (patching, certificate management, software releases) with safety, automation, and observability.
  • Design and operate production observability aligned to service health and customer impact.
  • Lead and participate in incident response, for high-severity events.
  • Collaborate with engineering, product, architecture, and operations teams.
  • Operate with autonomy and sound judgment in reliability decisions.
Qualifications & Requirements
  • 8+ years of hands-on Site Reliability Engineering or reliability-focused engineering experience with end-to-end service ownership.
  • Proven operation at a senior engineering scope with accountability for reliability outcomes.
  • Strong software engineering skills in C#, .NET, Java, Python, React, or similar technologies.
  • Practical experience applying SRE principles (SLIs, SLOs, error budgets).
  • Hands-on experience with AWS, Kubernetes, CI/CD, infrastructure as code and hybrid environments.
  • Strong knowledge of Linux and Windows systems, application platforms and relational databases.
  • Bachelor's or master's degree in computer science or equivalent experience.
  • Participation in anonymous-call rotation; flexible hours as required.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Site Reliability Engineer
Principal Site Reliability Engineer

Vertafore Career Center • Denver (CO)

On-site
USD 160,000 - 180,000
Director, Site Reliability Engineering
Director, Site Reliability Engineering

Vertafore Career Center • Denver (CO)

On-site
USD 175,000 - 220,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Vertafore • Denver (CO)

On-site
USD 110,000 - 145,000
Medical, vision & dental plans
401(k) Retirement Savings Plan with
Life, LTD/AD&D
+3
T Sr. Site Reliability Engineer Tiger Analytics Washington, Dc, US
T Sr. Site Reliability Engineer Tiger Analytics Washington, Dc, US

Artha Nexgen • Washington, Northern (KY)

Hybrid
USD 140,000 - 190,000
R Site Reliability Engineer Ii Restaurant365 Denver, Colorado, US
R Site Reliability Engineer Ii Restaurant365 Denver, Colorado, US

Artha Nexgen • Town of Texas (WI), Northern (KY)

Hybrid
USD 180,000 - 260,000
Senior SRE: Reliability, Observability & Auto-Healing
Senior SRE: Reliability, Observability & Auto-Healing

Vertafore Career Center • Denver (CO)

On-site
USD 110,000 - 145,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Senior SRE Architect: Enterprise Reliability & Scale
Senior SRE Architect: Enterprise Reliability & Scale

Vertafore • Denver (CO)

On-site
USD 160,000 - 180,000
Medical, Vision & Dental
401(k) + Employer Match
Parental Leave
+2
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Selby Jennings • Wilmington (NC)

On-site
USD 140,000 - 200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobgether • United States

Hybrid
USD 120,000 - 160,000
Competitive compensation package
Flexible work arrangements
Professional development opportunities
+2