Senior Site Reliability Engineer

Jobgether

India

Hybrid

INR 4,200,000 - 6,200,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work in Hyderabad
Health & life insurance

Job summary

Jobgether partner is seeking a Senior Site Reliability Engineer II based in India to design and operate secure, highly available cloud infrastructure for healthcare apps.

You will own AWS environments, lead Kubernetes operations, and advance observability, CI/CD, and IaC practices while mentoring engineers and promoting reliability across teams.

Qualifications

  • 9–12 years of professional experience in Site Reliability Engineering or cloud engineering with ownership of large AWS environments.
  • Strong hands-on AWS with EC2, S3, Lambda, and RDS; cost optimization and resource management.
  • Proven experience with Kubernetes, DataDog, Jenkins, and cloud-native operations.
  • Scripting in Python or Bash; IaC experience with Terraform; automation mindset.
  • Deep understanding of CI/CD, deployment automation, containerization, scalability, and production ops.
  • Experience implementing secure cloud ops in regulated environments (HIPAA, GDPR, SOC 2) is valuable.
  • Proven leadership: mentor engineers, drive initiatives, influence standards, communicate clearly.
  • Excellent written and verbal communication with stakeholders.

Responsibilities

  • Design and deploy secure, scalable AWS infrastructure focused on reliability and cost efficiency.
  • Manage and optimize EC2, S3, Lambda, and RDS resources.
  • Maintain monitoring and observability using DataDog for proactive issue detection.
  • Lead Jenkins-based deployment pipelines, automation, and release reliability.
  • Oversee Kubernetes clusters for scalability, resiliency, and consistent deployment.
  • Champion Infrastructure as Code with Terraform and Ansible to standardize environments.
  • Ensure security/compliance with HIPAA, GDPR, SOC 2 across cloud ops.
  • Lead technical initiatives, mentor engineers, and shape the roadmap.
  • Investigate incidents, identify risks, and architect long-lasting fixes.

Skills

Kubernetes
DataDog
Jenkins
Python
Bash
Terraform
Ansible
AWS
CI/CD
SRE

Tools

Kubernetes
DataDog
Jenkins
Terraform
Ansible

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer based in India.

As a Senior Site Reliability Engineer II, you will help architect and operate secure, highly available cloud infrastructure supporting business-critical healthcare applications.
You will take strategic ownership of AWS environments, driving reliability, scalability, performance, and cost optimization.
The role combines hands-on engineering with technical leadership across Kubernetes, CI/CD, observability, and infrastructure automation.
You will strengthen cloud operations through Infrastructure as Code, proactive monitoring, and resilient deployment practices.
Security and compliance are central, with responsibility for maintaining rigorous HIPAA, GDPR, and SOC 2 standards.
You will also mentor engineers, lead complex initiatives, and influence technical strategy across cross-functional teams.
This is an opportunity to make a direct impact on healthcare technology while working in a collaborative, innovation-focused environment.

Accountabilities:
  • Cloud Architecture & Reliability: Design, deploy, and continuously improve secure, scalable, and fault-tolerant AWS infrastructure, with a focus on availability, resilience, performance, and cost efficiency.
  • Infrastructure Operations: Manage and optimize AWS services including EC2, S3, Lambda, and RDS while improving resource utilization and operational efficiency.
  • Observability: Enhance and maintain monitoring and observability capabilities using DataDog, enabling proactive issue detection, performance analysis, and deep visibility across cloud environments.
  • CI/CD & Deployment: Lead the evolution of Jenkins-based deployment pipelines, improving automation, release reliability, deployment velocity, and engineering confidence.
  • Kubernetes: Manage and optimize containerized environments, improving scalability, consistency, resilience, and deployment practices.
  • Infrastructure as Code: Champion automation through Terraform, Ansible, and related technologies to reduce manual effort, standardize infrastructure, and improve operational efficiency.
  • Security & Compliance: Ensure cloud operations and infrastructure adhere to stringent security and regulatory requirements, including HIPAA, GDPR, and SOC 2.
  • Technical Leadership: Lead complex engineering initiatives, influence the technical roadmap, and make architecture decisions that strengthen long-term platform reliability.
  • Mentorship: Coach and mentor engineers, share technical expertise, encourage strong engineering practices, and foster a collaborative culture.
  • Problem Resolution & Continuous Improvement: Proactively identify reliability risks, investigate complex incidents, and architect durable solutions that prevent recurring operational issues.
Requirements:
  • Experience: 9–12 years of professional experience in Site Reliability Engineering, Cloud Engineering, or a closely related discipline, with demonstrated ownership of large-scale AWS environments.
  • AWS Expertise: Strong hands-on knowledge of AWS services such as EC2, S3, Lambda, and RDS, including experience with cost optimization and resource management.
  • SRE Tooling: Proven experience with Kubernetes, DataDog, Jenkins, and modern cloud-native operational practices.
  • Automation & Coding: Strong scripting capabilities in Python or Bash and professional experience with Infrastructure as Code tools, particularly Terraform.
  • Cloud & DevOps: Strong understanding of cloud architecture, deployment automation, CI/CD, containerization, scalability, availability, and production operations.
  • Security & Compliance: Experience implementing secure cloud operations and working with regulated environments or compliance frameworks is highly valuable.
  • Problem Solving: Strong analytical and troubleshooting skills, combined with an ownership mindset and the ability to design systems that proactively prevent failures.
  • Leadership: Demonstrated ability to lead technical initiatives, mentor engineers, influence engineering standards, and communicate effectively with technical and business stakeholders.
  • Communication: Excellent written and verbal communication skills, with the ability to translate complex technical concepts into clear objectives and recommendations.
  • Preferred Qualifications: AWS professional-level certifications, experience with serverless architectures or technologies such as Kafka, Kinesis, or Redshift, and previous healthcare technology or high-security environment experience are advantageous.
Benefits:
  • Hybrid working environment in Hyderabad, with flexibility designed to support effective ways of working.
  • Competitive benefits package designed to support health, well-being, and financial security.
  • Comprehensive health, accidental, and life insurance coverage, including family coverage.
  • Complimentary office lunches and dinners on select days, along with healthy workplace snacks.
  • Annual wellness allowance supporting employee well-being and productivity.
  • Earned, casual, and sick leave to support work-life balance.
  • Paid parental leave covering maternity, paternity, adoption, surrogacy, and abortion leave.
  • Bereavement and extended medical leave options.
  • Celebration leave and company-paid holidays.
  • Opportunity to influence cloud and SRE strategy at a senior technical level.
  • Meaningful impact on healthcare technology and systems that support improved patient outcomes.
  • Strong opportunities for technical leadership, innovation, mentorship, and professional growth.
  • Collaborative culture focused on engineering excellence, customer impact, and continuous improvement.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Employ Humanity • Bengaluru

Hybrid
INR 3,000,000 - 6,000,000
Flexible work scheduling
Paid time off
Comprehensive benefits
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

F-Prime Capital • Pune District

On-site
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SourcingXPress • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Senior Software Engineer- Site Reliability
Senior Software Engineer- Site Reliability

S&P Global • Bengaluru

On-site
INR 6,535,000 - 10,271,000
Health & Wellness
Flexible Downtime
Continuous Learning
+3
Senior Software Engineer- Site Reliability
Senior Software Engineer- Site Reliability

S&P Global • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Health care coverage
Generous time off
Access to career resources
+2
Senior DevOps / Site Reliability Engineer (SRE)
Senior DevOps / Site Reliability Engineer (SRE)

Aura Recruitment Solutions • Bengaluru

On-site
INR 450,000 - 750,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Hilabs • Pune District

On-site
INR 1,500,000 - 2,500,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Relatient • Pune District

On-site
INR 1,500,000 - 2,600,000
Life insurance
Accident coverage
Education reimbursement
+2
Sr. Site Reliability Engineer (Azure / GCP, Terraform, Python) — UnitedHealth Group (Optum) · H[...]
Sr. Site Reliability Engineer (Azure / GCP, Terraform, Python) — UnitedHealth Group (Optum) · H[...]

Cloudsoftsol • Hyderabad

Hybrid
INR 2,400,000 - 4,400,000
Comprehensive health coverage for self + family + parents
Sponsored Azure/GCP/Kubernetes certifications
Strong leave and parental-leave benefits