Sr. Site Reliability Engineer (SRE principles,SLIs, SLOs,AWSerror budgets, observability ,Debugging,Linux & Windows,C#, .NET)

Vertafore Career Center

Hyderabad

On-site

INR 1,200,000 - 1,600,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Vertafore is seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of our critical production services. This role covers the full-service lifecycle, from design and deployment readiness through production operations, incident response, and continuous improvement.

You will work across AWS, hybrid data centers, and customer-hosted environments, applying software engineering practices to automate resilience, monitor health, and

Responsibilities

  • Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health.
  • Define and manage SLIs and SLOs, using error budgets to guide delivery decisions.
  • Influence of service and system design to improve fault tolerance, observability and operational sustainability.
  • Debug complex production issues across application code, services and infrastructure using software engineering practices.
  • Perform root cause analysis using logs, metrics, traces, and code-level investigation.
  • Build automation and self-healing mechanisms to prevent repeat failures.
  • Execute production changes (patching, certificate management, software releases) with safety, automation, and observability.
  • Design and operate production observability aligned to service health and customer impact.
  • Lead and participate in incident response for high-severity events.
  • Collaborate with engineering, product, architecture, and operations teams.
  • Operate with autonomy and sound judgment in reliability decisions.

Job description

Vertafore is a leading technology company whose innovative software solution are advancing the insurance industry. Our suite of products provides solutions to our customers that help them better manage their business, boost their productivity and efficiencies, and lower costs while strengthening relationships.

Our mission is to move InsurTech forward by putting people at the heart of the industry. We are leading the way with product innovation, technology partnerships, and focusing on customer success.

Our fast-paced and collaborative environment inspires us to create, think, and challenge each other in ways that make our solutions and our teams better.

We are headquartered in Denver, Colorado, with offices across the U.S., Canada, and India.

We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is accountable for the full-service lifecycle, from design and deployment readiness through production operations, incident response, and continuous improvement. Reliability is a core engineering responsibility, requiring strong software engineering skills and autonomous operation across AWS, hybrid data centers, and customer-hosted environments.

Roles and Responsibilities
  • Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health.
  • Define and manage SLIs and SLOs, using error budgets to guide delivery decisions.
  • Influence of service and system design to improve fault tolerance, observability and operational sustainability.
  • Debug complex production issues across application code, services and infrastructure using software engineering practices.
  • Perform root cause analysis using logs, metrics, traces, and code-level investigation.
  • Build automation and self-healing mechanisms to prevent repeat failures.
  • Execute production changes (patching, certificate management, software releases) with safety, automation, and observability.
  • Design and operate production observability aligned to service health and customer impact.
  • Lead and participate in incident response for high-severity events.
  • Collaborate with engineering, product, architecture, and operations teams.
  • Operate with autonomy and sound judgment in reliability decisions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr.Manager, SRE (SRE, SLI/ SLO, monitoring, automation)
Sr.Manager, SRE (SRE, SLI/ SLO, monitoring, automation)

Vertafore Career Center • Hyderabad

On-site
INR 4,500,000 - 7,500,000
Sr Manager, Sre Hyderabad
Sr Manager, Sre Hyderabad

Vertafore • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Resilience and Reliability Engineer
Resilience and Reliability Engineer

EY • Pune District, Gurugram District, Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Infosys • Hyderabad

On-site
INR 1,400,000 - 2,200,000
Lead SRE
Lead SRE

Cvent • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Lead SRE
Lead SRE

Cvent, Inc. • India

On-site
INR 2,500,000 - 4,500,000
Associate Site Reliability Engineer II
Associate Site Reliability Engineer II

MetLife • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer
Site Reliability Engineer

Metlife • Hyderabad

Hybrid
INR 1,500,000 - 2,600,000
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000