Sr. Site Reliability Engineer (Azure / GCP, Terraform, Python) — UnitedHealth Group (Optum) · Hyderabad

Cloud Soft Solutions

Hyderabad

Hybrid

INR 2,400,000 - 4,400,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health coverage and benefits
Hybrid work model
Certifications learning budget

Job summary

Cloud Soft Solutions is seeking a Senior SRE in Hyderabad to build reliable Azure and GCP platforms with Terraform and Python. You will run AKS/GKE clusters, implement SLO-driven reliability, and collaborate on multi-region disaster recovery in a healthcare context.

Ideal candidates have 5+ years in SRE/DevOps with Terraform expertise, Kubernetes experience, and strong Python/CI-CD skills. The role offers a hybrid work arrangement in Hyderabad and a competitive INR salary package.

Qualifications

  • 5+ years of SRE / DevOps / Platform Engineering experience with a strong reliability mindset.
  • Terraform expertise on Azure and / or GCP — reusable modules, remote state, plan reviews.
  • Production Kubernetes experience (AKS or GKE preferred).
  • Strong Python for automation tooling and CI/CD; comfortable with Bash.

Responsibilities

  • Build and maintain reliable cloud platforms using Terraform for infrastructure-as-code on Azure and GCP.
  • Manage Kubernetes clusters (AKS and GKE): cluster lifecycle and upgrades, autoscaling, IAM, ingress, and Helm-based add-ons.
  • Implement observability with Prometheus/Grafana and automated remediation to reduce toil.
  • Integrate Python scripting for reliability tooling, CI/CD optimisation, and runbooks.
  • Ensure HIPAA-aware security, audit trails, and FinOps reporting in a regulated healthcare environment.
  • Collaborate cross-functionally on disaster-recovery strategies including multi-region failover.

Skills

Terraform
Azure/GCP
Kubernetes
Python
Bash
Prometheus
Grafana
CI/CD
SLI/SLO
Security/Compliance
FinOps

Education

Bachelor's in CS / IT or equivalent

Job description

Senior SRE for UnitedHealth Group / Optum's healthcare-scale platforms in Hyderabad — build reliable Azure / GCP infrastructure with Terraform and Python, run AKS / GKE clusters, and drive SLO-led reliability engineering. 5+ years, hybrid Hyderabad, INR 24-44 LPA band.

Responsibilities

Build and maintain reliable cloud platforms using Terraform for infrastructure-as-code on Azure and GCP, supporting healthcare-scale systems where uptime directly affects patients and providers.Manage Kubernetes clusters (AKS and / or GKE): cluster lifecycle and upgrades, autoscaling, Workload Identity for least-privilege IAM, ingress, and platform add-ons via Helm.Implement observability, alerting and automated remediation — define and track SLIs / SLOs, instrument services with Prometheus / Grafana and cloud-native monitoring, and reduce toil through self-healing automation.Integrate Python scripting for custom reliability tooling, CI/CD optimisation, and infrastructure management; conduct blameless post-mortems and codify learnings into runbooks.Ensure compliance, security and cost governance in a regulated healthcare environment — secrets management, encryption, audit trails (HIPAA-aware), and FinOps reporting.Collaborate cross-functionally to deliver resilient microservices and disaster-recovery strategies, including multi-region failover and tested restore procedures.

Requirements

5+ years of SRE / DevOps / Platform Engineering experience with a strong reliability mindset.Terraform expertise on Azure and / or GCP — reusable modules, remote state, plan reviews.Production Kubernetes experience (AKS or GKE preferred).Strong Python for automation tooling and CI/CD; comfortable with Bash.Hands-on observability (Prometheus, Grafana, cloud-native monitoring) and incident-response tooling.SLI / SLO definition, error budgets, and post-incident review discipline.Experience in regulated or large-scale environments (healthcare, BFSI) is a strong plus; HIPAA awareness valued.Bachelor's in CS / IT or equivalent. Azure / GCP and CKA certifications a plus.

Why this role in 2026

Optum (the technology and health-services arm of UnitedHealth Group, a Fortune-5 company) runs one of the largest healthcare-technology engineering centres in Hyderabad. Reliability and compliance are first-class priorities here, which makes it an excellent environment to deepen genuine SRE skills — SLOs, error budgets, multi-cloud Terraform and disaster recovery — inside a stable, well-funded organisation.

Company & benefits

INR 24-44 LPA fixed band for senior profiles plus annual bonus (market-rate estimate; confirmed at offer). Comprehensive health coverage for self + family + parents. Hybrid model in Hyderabad. Sponsored Azure / GCP / Kubernetes certifications and learning budget. Strong leave, parental-leave and wellness benefits typical of a Fortune-5 employer.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer III (Kubernetes, Terraform, Multi-Cloud) — Amgen · Hyderabad
Site Reliability Engineer III (Kubernetes, Terraform, Multi-Cloud) — Amgen · Hyderabad

Cloud Soft Solutions • Hyderabad

Hybrid
INR 2,600,000 - 4,600,000
Medical coverage
Hybrid work
Certifications sponsorship
+1
Site Reliability Engineer II — Performance Engineer (Kubernetes, Terraform, Azure / AWS) — Quest Diagnostics · Hyderabad
Site Reliability Engineer II — Performance Engineer (Kubernetes, Terraform, Azure / AWS) — Quest Diagnostics · Hyderabad

Cloud Soft Solutions • Hyderabad

Hybrid
INR 2,000,000 - 3,800,000
Hybrid setup in Hyderabad
Comprehensive health benefits
Senior Site Reliability Engineer
Senior Site Reliability Engineer

UnitedHealth Group • Hyderabad

On-site
Confidential
SRE Production DevOps Engineer (EKS, Terraform, Ansible, CI/CD) — GE Vernova · Hyderabad
SRE Production DevOps Engineer (EKS, Terraform, Ansible, CI/CD) — GE Vernova · Hyderabad

Cloud Soft Solutions • Hyderabad

Hybrid
INR 2,600,000 - 4,800,000
Comprehensive medical
Stock / ESPP eligibility
Hybrid working in Hyderabad
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

UnitedHealth Group • Dadri

On-site
INR 2,500,000 - 4,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SourcingXPress • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Optum India • Hyderabad

On-site
INR 1,200,000 - 2,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Taleo • Dadri

On-site
INR 2,500,000 - 4,500,000
Site Reliability Engineer
Site Reliability Engineer

Acesoft Labs • Hyderabad

On-site
INR 1,800,000 - 3,000,000
DevOps Engineering Expert (AWS, Kubernetes, MLOps) — Sanofi · Hyderabad
DevOps Engineering Expert (AWS, Kubernetes, MLOps) — Sanofi · Hyderabad

Cloud Soft Solutions • Hyderabad

Hybrid
INR 2,400,000 - 4,200,000
Hybrid work
Medical cover
Stock plan eligibility
+2