Senior Site Reliability Engineer

Optum

Hyderabad

On-site

INR 4,000,000 - 7,000,000

Full time

29 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Optum seeks a Senior Site Reliability Engineer to design, deploy, and manage Azure infrastructure and Kubernetes clusters in Hyderabad. You will build scalable, secure, and resilient platforms and create CI/CD pipelines for data teams.

You will implement IaC with Terraform/Ansible, monitor with Grafana/Prometheus, and collaborate across teams to drive reliability and automation while mentoring juniors and staying current with DevOps trends.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field or equivalent work experience.
  • 5+ years of experience in DevOps, SRE, or similar role.
  • Experience with CI/CD tools, preferably GitHub Actions.
  • Solid experience with Data and AI platforms (Databricks, Snowflake).
  • Experience with orchestration tools (Airflow, Data Factory).
  • Infrastructure as Code experience (Terraform, CloudFormation).
  • Expert knowledge of Azure and core Azure services (Storage, Networking, AKS).
  • Scripting skills (Python, Bash).
  • Monitoring tools experience (Grafana, Prometheus, Splunk).
  • Proven Kubernetes and containerization expertise (Docker, Podman).
  • Strong problem-solving, communication, and collaboration skills.

Responsibilities

  • Design, deploy, and manage Azure infrastructure and services.
  • Optimize cloud resource utilization and cost in Azure.
  • Implement IaC using Terraform/Ansible to provision resources.
  • Build and manage Kubernetes clusters for containerized apps.
  • Develop CI/CD pipelines for data engineering and data science teams.
  • Create automation and self-service capabilities for researchers.
  • Ensure platform security, performance, and high availability.
  • Mentor junior engineers and promote DevOps best practices.
  • Implement observability with monitoring and logging solutions.
  • Stay current with industry trends and drive process improvements.
  • Develop AI-powered solutions using no-code/low-code platforms.

Skills

Kubernetes
Docker
CI/CD

Education

Bachelor's degree in CS/Engineering

Tools

Terraform
GitHub Actions
Ansible
CloudFormation
Airflow
Databricks
Snowflake
Grafana
Prometheus
AKS
Python
Bash
Podman

Job description

Improve the lives of others while Caring. Connecting. Growing together.

Job Description - Senior Site Reliability Engineer (2385843)

Senior Site Reliability Engineer - 2385843

Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, data and resources they need to feel their best. Here, you will find a culture guided by inclusion, talented peers, comprehensive benefits and career development opportunities. Come make an impact on the communities we serve as you help us advance health optimization on a global scale. Join us to start Caring. Connecting. Growing together.

Primary Responsibilities:
  • Design, deploy, and manage Azure infrastructure and services
  • Optimize cloud resource utilization and cost management in Azure.
  • Infrastructure as Code (IaC):
    • Utilize IaC tools (such as Terraform, Ansible, or similar) to provision and manage infrastructure
    • Ensure infrastructure is scalable, secure, and resilient
  • Design, deploy, and manage Kubernetes clusters to support containerized applications
  • Implement and manage Kubernetes-based solutions for orchestration, scaling, and security
  • AI/DevOps/MLOps:
    • Design, build, and maintain robust, automated CI/CD pipelines for the data engineering and data science teams
    • Develop and manage tools that support these teams, ensuring a seamless and efficient experience for them
    • Utilize and build AI solutions to drive efficiencies across the platform engineering and wider teams
  • Automation & Self-Service:
    • Implement automation for various operational processes, reducing manual intervention
    • Create self-service capabilities that empower Data Scientists to deploy and manage their applications independently
  • Platform Security & Performance:
    • Ensure the security of the platform by implementing best practices and monitoring for vulnerabilities
    • Continuously monitor and optimize the performance of the platform to ensure high availability and reliability
  • Collaborate with cross-functional teams to align on project requirements and deliverables
  • Mentor junior team members, promoting best practices in DevOps and automation
  • Observability:
    • Implement monitoring and logging solutions to ensure system health and performance
    • Troubleshoot and resolve issues related to system performance, security, and reliability
  • Stay current with industry trends and advancements in DevOps practices and technologies
  • Identify opportunities for process improvements and drive initiatives to implement them
  • Design, develop, and deploy AI-powered solutions using no-code, low-code, and advanced platforms, translating business needs into scalable applications that enhance products, workflows, and decision-making
  • Comply with the terms and conditions of the employment contract, company policies and procedures, and any and all directives (such as, but not limited to, transfer and/or re-assignment to different work locations, change in teams and/or work shifts, policies in regards to flexibility of work benefits and/or work environment, alternative work arrangements, and other decisions that may arise due to the changing business environment). The Company may adopt, vary or rescind these policies and directives in its absolute discretion and without any limitation (implied or otherwise) on its ability to do so
Required Qualifications:
  • Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent work experience)
  • 5+ years of experience in DevOps, Site Reliability Engineering (SRE), or a similar role
  • Experience with CI/CD tools preferably GitHub Actions
  • Solid experience of Data and AI platforms, preferably Databricks and Snowflake
  • Experience using orchestrating tools (Airflow, Data Factory)
  • Experience with Infrastructure as Code (Terraform, CloudFormation)
  • Expert knowledge of cloud platforms, with a focus on Azure
  • Familiarity with core Azure services - Storage, Networking, security, App Services, AKS
  • Proficiency in scripting languages (e.g., Python, Bash, Ruby)
  • Proficiency in monitoring tools (Splunk, Grafana, Prometheus)
  • Proven expertise in Kubernetes and containerization technologies (Docker, Podman)
  • Proven excellent problem-solving and analytical skills
  • Proven solid communication and collaboration skills
  • Proven ability to work independently and in a team-oriented, collaborative environment
  • Proven ability to work with third party vendors and support teams in a large multi disciplined organization
  • Skills:
    • Kubernetes, Docker and Containerization, CICD
    • Azure Cloud
    • Infrastructure as Code
    • Grafana and Prometheus
    • AKS
Preferred Qualifications:
  • Experience in Azure, AWS, GCP and private cloud technologies
  • Experience using Firewalls, Load Balancers and DDOS solutions
  • Experience in a leadership or mentorship role
  • Knowledge of security best practices in DevOps

At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.

UnitedHealth Group is committed to working with and providing reasonable accommodations to individuals with physical and mental disabilities. If you need special assistance or accommodation for any part of the application process, please call 1-866-566-8715 to be connected to Recruitment Services. Recruitment Services hours of operation are 7 a.m. to 7 p.m. CT, Monday through Friday.

UnitedHealth Group is a registered service mark of UnitedHealth Group, Inc. The UnitedHealth Group name with the dimensional logo, as well as the dimensional logo alone, are both service marks for the UnitedHealth Group, Inc.

Diversity creates a healthier atmosphere: UnitedHealth Group is an Equal Employment Opportunity/Affirmative Action employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, age, national origin, protected veteran status, disability status, sexual orientation, gender identity or expression, marital status, genetic information, or any other characteristic protected by law.

UnitedHealth Group is a drug-free workplace. Candidates are required to pass a drug test before beginning employment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

UnitedHealth Group • Hyderabad

On-site
Confidential
Senior I O Engineering Consultant - Python, K8s, Azure cloud, Terraform
Senior I O Engineering Consultant - Python, K8s, Azure cloud, Terraform

Optum India • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Devops Engineer - SRE, Cloud Engineering
Devops Engineer - SRE, Cloud Engineering

Optum • Hyderabad

On-site
INR 2,500,000 - 4,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Optum India • Chennai District

On-site
INR 350,000 - 750,000
Senior Data Analyst
Senior Data Analyst

Optum • Dadri

On-site
INR 1,800,000 - 3,200,000
Senior Platform Engineering Consultant
Senior Platform Engineering Consultant

Optum • Hyderabad

On-site
INR 2,500,000 - 5,000,000
Senior Software Engineer
Senior Software Engineer

Optum • Hyderabad

On-site
INR 1,800,000 - 2,400,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Optum • Chennai District

On-site
INR 1,400,000 - 2,500,000
Software Engineer Lead
Software Engineer Lead

Optum • Hyderabad

On-site
INR 2,500,000 - 4,500,000
Senior Software Engineer I - Azure DevOps
Senior Software Engineer I - Azure DevOps

Optum India • Bengaluru

On-site
INR 3,000,000 - 5,400,000