Senior SRE: Reliability, Observability & Auto-Healing

Vertafore Career Center

Denver (CO)

On-site

USD 110,000 - 145,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Vertafore is seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services in Denver, CO. You will oversee the full-service lifecycle—from design and deployment readiness through production operations and incident response.

This role emphasizes strong software engineering, autonomous operation across AWS, hybrid data centers, and customer-hosted environments, with a focus on observability, fault tolerance,

Qualifications

  • 8+ years of hands-on Site Reliability Engineering or reliability-focused engineering experience with end-to-end service ownership.
  • Senior engineering scope with accountability for reliability outcomes.
  • Strong software engineering skills in C#, .NET, Java, Python, React, or similar technologies.
  • Practical experience applying SRE principles (SLIs, SLOs, error budgets).
  • Hands-on experience with AWS, Kubernetes, CI/CD, infrastructure as code and hybrid environments.
  • Strong knowledge of Linux and Windows systems, application platforms and relational databases.
  • Bachelor's or master's degree in computer science or equivalent experience.
  • Participation in anonymous-call rotation; flexible hours as required.

Responsibilities

  • Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health.
  • Define and manage SLIs and SLOs, using error budgets to guide delivery decisions.
  • Influence of service and system design to improve fault tolerance, observability and operational sustainability.
  • Debug complex production issues across application code, services and infrastructure using software engineering practices.
  • Perform root cause analysis using logs, metrics, traces, and code-level investigation.
  • Build automation and self-healing mechanisms to prevent repeat failures.
  • Execute production changes (patching, certificate management, software releases) with safety, automation, and observability.
  • Design and operate production observability aligned to service health and customer impact.
  • Lead and participate in incident response, for high-severity events.
  • Collaborate with engineering, product, architecture, and operations teams.
  • Operate with autonomy and sound judgment in reliability decisions.

Skills

C#
Java
Python
.NET
React
SRE principles
Observability
Incident response
AWS
Kubernetes
CI/CD
Linux/Windows

Education

Bachelor's or Master's in CS or equivalent

Tools

Terraform
CI/CD pipelines
AWS Console
Git

Job description

Vertafore is seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services in Denver, CO. You will oversee the full-service lifecycle—from design and deployment readiness through production operations and incident response.

This role emphasizes strong software engineering, autonomous operation across AWS, hybrid data centers, and customer-hosted environments, with a focus on observability, fault tolerance,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer: Scale & Observability
Senior Site Reliability Engineer: Scale & Observability

Vertafore • Denver (CO)

On-site
USD 110,000 - 145,000
Medical, vision & dental plans
401(k) Retirement Savings Plan with
Life, LTD/AD&D
+3
Senior SRE Architect: Enterprise Reliability & Scale
Senior SRE Architect: Enterprise Reliability & Scale

Vertafore • Denver (CO)

On-site
USD 160,000 - 180,000
Medical, Vision & Dental
401(k) + Employer Match
Parental Leave
+2
Enterprise SRE Architect – Reliability & Observability
Enterprise SRE Architect – Reliability & Observability

Vertafore Career Center • Denver (CO)

On-site
USD 160,000 - 180,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Vertafore Career Center • Denver (CO)

On-site
USD 110,000 - 145,000
Director, Reliability & Observability Engineering
Director, Reliability & Observability Engineering

Vertafore Career Center • Denver (CO)

On-site
USD 175,000 - 220,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Vertafore Career Center • Denver (CO)

On-site
USD 160,000 - 180,000
Director, Site Reliability Engineering
Director, Site Reliability Engineering

Vertafore Career Center • Denver (CO)

On-site
USD 175,000 - 220,000
Senior SRE - Cloud Reliability & Automation
Senior SRE - Cloud Reliability & Automation

Frontier-Airlines • Denver (CO)

Hybrid
USD 110,000 - 146,000
Medical/Dental/Vision
401(k) retirement plan
Travel privileges
+3
Senior SRE - Cloud & Observability
Senior SRE - Cloud & Observability

Ridgeline • Reno (NV)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Education reimbursement
Wellness reimbursement
+1
Remote Senior SRE — Cloud Reliability & Automation
Remote Senior SRE — Cloud Reliability & Automation

EverCommerce • Denver (CO)

Hybrid
USD 110,000 - 130,000
Flexible work environment
Health and wellness benefits
401(k) with company match
+2