Sr. Site Reliability Engineer

Relatient

Pune District

On-site

INR 1,500,000 - 2,600,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Life insurance
Accident coverage
Education reimbursement
Public holidays & floating day
Hybrid work

Job summary

Relatient is seeking a Senior Site Reliability Engineer to lead reliability, observability, and operational excellence across our cloud-hosted platform. You will own production readiness, manage major incidents, and drive automation to prevent customer impact.

The role requires 8+ years in enterprise SaaS or cloud-native environments, hands-on cloud, monitoring, and incident management experience, and collaboration with multiple teams to improve resilience. Hybrid work in India is offered.

Qualifications

  • Bachelor's degree in computer science, engineering, or a related field.
  • 8+ years supporting enterprise SaaS or cloud-native production environments.
  • Experience leading Sev1/Sev2 incidents and post-incident reviews.
  • Hands-on in cloud infrastructure, monitoring, observability, and incident management.
  • Experience in 24x7 production support environments.

Responsibilities

  • Own operational excellence and reliability of Relatient's platforms and services.
  • Lead P1/P2 incidents with coordination, communications, and post-incident reviews.
  • Detect and resolve production issues through monitoring and observability.
  • Design and improve SLIs/SLOs, error budgets, and KPIs.
  • Monitor cloud-native apps, APIs, databases, messaging platforms, and integrations.
  • Review releases for readiness and post-deployment verification.
  • Collaborate with cross-functional teams to improve resilience and scalability.
  • Mentor engineers and promote best practices in SRE.

Skills

Observability
Incident management
Leadership
DevOps practices
Monitoring
Communication

Education

Bachelor's degree in CS / Engineering

Tools

AWS
Kubernetes
Docker

Job description

At Relatient, we help healthcare organizations optimize patient access through AI-powered workflows, real-time automation, and flexible access tools. We are trusted by over 50,000 providers to modernize the patient experience and have been recognized by Forbes and Deloitte for our innovative and inclusive culture.

Your Role at Relatient:

We’re seeking a Senior Site Reliability Engineer (SRE) to join our team to lead production reliability, observability, and operational excellence across our cloud-hosted healthcare platform. Our office is in Amar Tech Park and brings in an amazing culture where we focus on work that makes a difference.

This role is responsible for ensuring the availability, performance, scalability, and resiliency of mission‑critical customer‑facing platforms, including Patient Scheduling, Patient Engagement, Voice, Chat, Messaging, APIs, Reporting, and Healthcare Integrations.

The Senior SRE partners closely with Engineering, Infrastructure, Security, Product, Customer Support, and Implementation teams to proactively prevent customer‑impacting incidents, improve production stability, automate operational processes, and continuously enhance the reliability of our cloud‑native applications and services.

How You’ll Make an Impact:
  • Own the operational excellence, reliability, availability, performance, and production readiness of Relatient's customer‑facing platforms and services.
  • Lead P1/P2 production incidents, coordinating end‑to‑end incident response, stakeholder communication, service restoration, root cause analysis, and post‑incident reviews.
  • Proactively detect, investigate, and resolve production issues before they impact customers through effective monitoring, observability, and operational reviews.
  • Design, implement, and continuously improve observability, monitoring, alerting, SLIs, SLOs, Error Budgets, and operational KPIs.
  • Monitor and troubleshoot cloud‑native applications, APIs, databases, messaging platforms, third‑party integrations, and production workloads.
  • Ensure the health and availability of customer‑facing services including Patient Scheduling, Voice, Chat, SMS, Email, Reporting, Healthcare Integrations, and EHR platforms.
  • Support production cloud infrastructure, deployments, configurations, platform services, and operational readiness across production environments.
  • Review and validate production releases through deployment readiness reviews, health checks, smoke testing, dependency validation, and post‑deployment verification.
  • Participate in architecture reviews, release planning, change management, and Go/No‑Go decisions to ensure production stability and reliability.
  • Collaborate with Engineering, Infrastructure, Security, Product, Customer Support, and third‑party vendors to improve platform resilience, scalability, and customer experience.
  • Drive automation, monitoring optimization, configuration governance, and continuous operational improvements.
  • Develop and maintain operational runbooks, incident playbooks, recovery procedures, and technical documentation.
  • Conduct blameless postmortems and implement corrective and preventive actions to eliminate recurring operational issues.
  • Mentor Site Reliability Engineers while promoting operational standards, engineering best practices, and a culture of proactive ownership and continuous improvement.
What You Bring:
  • Bachelor’s degree in computer science, B.E./ B. Tech, computer engineering, or a related technical discipline.
  • 8+ years supporting enterprise SaaS or cloud‑native production environments.
  • Experience supporting large‑scale production applications and customer‑facing platforms.
  • Hands‑on experience with cloud infrastructure, application monitoring, observability, and incident management.
  • Experience leading Sev1/Sev2 incidents, root cause analysis, and operational reviews.
  • Experience working within 24x7 production support environments.
  • Experience supporting one or more of the following:
    • Modern programming languages and application frameworks (Java, Spring Boot, PHP, Node.js, Angular, React)
    • Cloud platforms and infrastructure services (AWS preferred)
    • APIs, distributed systems, messaging platforms, and third‑party integrations
    • SQL and NoSQL databases, caching, and messaging technologies
    • Observability, monitoring, logging, and distributed tracing platforms
    • CI/CD, source control, containerization, automation, and DevOps practices
    • Linux/Unix operating systems and scripting
    • ITIL Incident, Problem, and Change Management processes
Mindsets That Matter:

We always look for ways to grow and take pride in what we do. You'll thrive here if you:

  • Act with purpose, focus, and accountability
  • Collaborate across teams and communicate clearly
  • Keep innovating and automate what slows you down
Benefits Include (India):
  • INR 5,00,000life insurance for all full‑time employees and their immediate household.
  • INR 15,00,000accident coverage
  • Education reimbursement
  • 10 national/state holidays and 1 floating holiday
  • Flexible hours and hybrid work.
Equal Opportunity at Relatient:

We’re building a team as diverse as the communities we serve. Relatient is proud to be an equal opportunity employer. If you need accommodation during the application process, just let us know.

To learn more about our organization, visit www.relatient.com

Ready to Join Relatient?

If you’re looking for work that matters and a team that makes it count, we'd love to hear from you!

#LI-AM1
#LI-Hybrid

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Software Engineer
Lead Software Engineer

Relatient • Pune District

On-site
INR 1,500,000 - 2,500,000
Life insurance
Accident coverage
Education reimbursement
+3
Lead Software Engineer Hybrid Remote, Pune, Maharashtra Lead Software Engineer
Lead Software Engineer Hybrid Remote, Pune, Maharashtra Lead Software Engineer

Relatient • Maharashtra

Hybrid
INR 2,000,000 - 3,800,000
Life insurance
Accident coverage
Education reimbursement
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SourcingXPress • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Implementation Analyst
Implementation Analyst

Relatient • Pune District

On-site
INR 600,000 - 1,200,000
Life Insurance
Accident Coverage
Education Reimbursement
+2
Site Reliability Engineer
Site Reliability Engineer

CrelioHealth INC • Pune District

On-site
INR 1,500,000 - 2,500,000
Open & flexible culture
Youthful team atmosphere
Opportunity to innovate in healthcare
Lead Data Engineer
Lead Data Engineer

Relatient • Pune District

Hybrid
INR 2,500,000 - 4,000,000
Life insurance (INR 5,00,000)
Accident coverage (INR 15,00,000)
Education reimbursement
+2
Senior Site Reliability Engineer - Digital Health Products
Senior Site Reliability Engineer - Digital Health Products

Roche • Pune District

On-site
INR 1,500,000 - 2,200,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

United States Digital Space LLC • Bengaluru

Hybrid
INR 6,000,000 - 12,000,000
Site Reliability Engineer
Site Reliability Engineer

Persistent Systems • Pune District

Hybrid
INR 1,200,000 - 2,200,000
Hybrid work
Career growth
Education sponsorship
+4
Lead SRE
Lead SRE

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,400,000