Site Reliability Engineer - Level - 3

CorroHealth Infotech Private Limited

Dadri

On-site

INR 1,500,000 - 2,200,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

CorroHealth Infotech Private Limited in India is seeking a skilled Site Reliability Engineer (SRE) to ensure reliability, scalability, and performance across cloud platforms. You will design and run cloud infrastructure on AWS and Azure, manage Kubernetes clusters, implement IaC, and drive automation with Python and CI/CD.

The role requires collaboration with Dev, QA, and Security teams to maintain production readiness.

Qualifications

  • 3–6 years in SRE, DevOps, or Cloud Engineering.
  • Strong experience with AWS (EC2, S3, Lambda, CloudWatch, RDS, VPC) and Azure (VMs, Functions, Monitor, Networking).
  • Hands-on with Kubernetes (EKS, AKS) and Docker.
  • Proficient in Python for automation and system integrations.
  • Experience with CI/CD tools: Jenkins, GitHub Actions, GitLab, Azure DevOps.
  • Proficiency in IaC: Terraform / CloudFormation / Bicep.
  • Observability: Prometheus, Grafana, ELK/EFK, Datadog, New Relic.

Responsibilities

  • Design, deploy, and optimize workloads on AWS and Azure.
  • Manage Kubernetes clusters (EKS, AKS) and serverless (Lambda).
  • Implement secure, scalable infrastructure with cost-optimization.
  • Build automation frameworks and tooling with Python.
  • Implement IaC using Terraform, CloudFormation, or Bicep.
  • Automate deployments, scaling, monitoring, and remediation.
  • Develop and maintain CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI/CD, Azure DevOps).
  • Collaborate with Dev, QA, and Security teams to ensure production readiness.
  • Establish end-to-end observability: metrics, logs, traces.

Skills

Analytical thinking
Troubleshooting
Collaboration
Strong communication

Tools

AWS
Azure
Kubernetes
Docker
Terraform
CloudFormation
Bicep
Jenkins
GitHub Actions
GitLab CI/CD
Azure DevOps
Prometheus
Grafana
ELK/EFK
Datadog
New Relic

Job description

About Us

Our purpose is to help clients exceed their financial health goals. Across the reimbursement cycle, our scalable solutions and clinical expertise help solve programmatic needs. Enabling our teams with leading technology allows analytics to guide our solutions and keeps us accountable achieving goals. We build long-term careers by investing in YOU. We seek to create an environment that cultivates your professional development and personal growth, as we believe your success is our success.

ESSENTIAL DUTIES AND RESPONSIBILITIES

Note: The essential duties and responsibilities below are intended to describe the general duties and responsibilities of this position and are not intended to be an exhaustive statement of duties. This position may perform all or most of the primary duties listed below. Specific tasks, responsibilities or competencies may be documented in the Team Member’s performance objectives as outlined by the Team Member’s immediate Leadership Team Member. We are seeking a skilled Site Reliability Engineer (SRE) with expertise in AWS, Azure, Kubernetes, Serverless (Lambda), Python, CI/CD, and observability tooling. The SRE will ensure system reliability, scalability, and performance across our cloud platforms, while driving automation, observability, and operational excellence.

Key Responsibilities
  • Cloud & InfrastructureDesign, deploy, and optimize workloads on AWS and Azure.
  • Manage Kubernetes clusters (AKS, EKS) and serverless solutions like AWS Lambda.
  • Implement secure, scalable, and cost-optimized infrastructure.
  • Automation & EngineeringBuild automation frameworks and operational tooling using Python.
  • Implement Infrastructure as Code (IaC) using Terraform, CloudFormation, or Azure Bicep.
  • Automate deployments, scaling, monitoring, and remediation.
  • CI/CD & DevOps PracticesDevelop and maintain CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI/CD, Azure DevOps).
  • Partner with development teams to improve release velocity and reliability.
  • Observability & MonitoringImplement end-to-end observability practices: metrics, logging, tracing.
  • Manage logging & monitoring solutions (Loggly,ELK/EFK stack, CloudWatch, Azure Monitor, Prometheus, Grafana, Datadog, New Relic).
  • Establish alerting & escalation workflows with PagerDuty, OpsGenie, or similar incident management platforms.
  • Site Reliability & Incident ManagementParticipate in 24/7 on-call rotations for critical systems.
  • Diagnose, resolve, and perform root cause analysis for incidents.
  • Drive post-incident reviews and continuous improvement in reliability.
  • Collaboration & MentorshipWork closely with Dev, QA, and Security teams to ensure production readiness.
  • Champion SRE best practices across teams.
  • Mentor junior engineers on cloud-native operations and observability.
Required Skills & Qualifications
  • Experience: 3–6 years in SRE, DevOps, or Cloud Engineering.
  • Cloud Platforms: Strong experience with AWS (EC2, S3, Lambda, CloudWatch, RDS, VPC, etc.) and Azure (VMs, Functions, Monitor, Networking, etc.).
  • Containers & Orchestration: Hands-on with Kubernetes (EKS, AKS) and containerization (Docker).
  • Programming: Proficient in Python for automation and system integrations.
  • CI/CD Tools: Experience with Jenkins, GitHub Actions, GitLab, Azure DevOps.
  • IaC: Proficiency with Terraform / CloudFormation / Bicep.
  • Observability: Experience with Prometheus, Grafana, ELK/EFK, Datadog, New Relic, or similar.
  • Incident Management: Familiar with PagerDuty, OpsGenie, or equivalent tools.
  • Networking & Security: Solid understanding of IAM, VPC, firewalls, encryption, and compliance.
  • Soft Skills: Strong analytical, troubleshooting, and collaboration skills.
  • Preferred Qualifications Certifications: AWS Certified SysOps Administrator / Solutions Architect, Microsoft Azure Administrator Associate, CKA (Certified Kubernetes Administrator).
  • Experience with multi-cloud deployments.
  • Familiarity with service mesh (Istio/Linkerd) and advanced observability (OpenTelemetry)
PHYSICAL DEMANDS

Note: Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions as described. Regular eye-hand coordination and manual dexterity is required to operate office equipment. The ability to perform work at a computer terminal for 6-8 hours a day and function in an environment with constant interruptions is required. At times, Team Member are subject to sitting for prolonged periods. Infrequently, Team Member must be able to lift and move material weighing up to 20 lbs. Team Member may experience elevated levels of stress during periods of increased activity and with work entailing multiple deadlines.

A job description is only intended as a guideline and is only part of the Team Member’s function. The company has reviewed this job description to ensure that the essential functions and basic duties have been included. It is not intended to be construed as an exhaustive list of all functions, responsibilities, skills and abilities. Additional functions and requirements may be assigned by supervisors as deemed appropriate. CorroHealth sits at the center of the revenue cycle revolution. Fundamental operations of the revenue cycle are supported through our expert teams while we recast the role of clinicians through automation. This shift to a true clinical revenue cycle helps us achieve our core purpose – exceed client financial health goals. For each patient population, CorroHealth automates key clinical aspects of the cycle. Our platforms focus on capture and application of clinical documentation while easing the burden on physicians. Scalability is prioritized in the support of client program operations. As with most revenue cycle partners, our skilled and enthusiastic team is available to outsource any portion of the cycle. However, we can also complement client programs with additional expert support or upskill existing client teams to meet program demands. Whether our team is deployed directly, or automation is incorporated for a more programmatic solution, CorroHealth delivers. CorroHealth has acquired Xtend Healthcare! For more information, please visit https://corrohealth.com. Applicants will only receive job-related emails from the domain @corrohealth.com. Additionally, it is important to emphasize that CorroHealth will never ask for money in return for a job offer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AM - RCM Services
AM - RCM Services

CorroHealth Infotech Private Limited • Hyderabad

On-site
INR 600,000 - 900,000
Jr Executive - HIM Services
Jr Executive - HIM Services

CorroHealth Infotech Private Limited • Dadri

On-site
INR 250,000 - 450,000
Sr Executive - MIS
Sr Executive - MIS

CorroHealth Infotech Private Limited • Dadri

On-site
INR 600,000 - 900,000
AVP - L&D
AVP - L&D

CorroHealth Infotech Private Limited • Bengaluru

On-site
INR 3,500,000 - 6,500,000
Sr Executive - Finance
Sr Executive - Finance

CorroHealth Infotech Private Limited • Chennai District

On-site
INR 450,000 - 650,000
Director – Security & Government affairs
Director – Security & Government affairs

CorroHealth Infotech Private Limited • Dadri

On-site
INR 3,000,000 - 6,000,000
QA - Software Testing
QA - Software Testing

CorroHealth Infotech Private Limited • Dadri

On-site
INR 600,000 - 1,200,000
Manager - Finance
Manager - Finance

CorroHealth Infotech Private Limited • Chennai District

On-site
INR 1,800,000 - 3,000,000
Lead - Application Support Engineer
Lead - Application Support Engineer

CorroHealth Infotech Private Limited • Hyderabad

On-site
INR 1,000,000 - 1,500,000
Manager - Quality
Manager - Quality

CorroHealth Infotech Private Limited • Dadri

On-site
INR 2,800,000 - 4,200,000