Senior Site Reliability Engineer (SRE)

JobCubby

India

On-site

INR 2,500,000 - 5,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hewlett Packard Enterprise (HPE) seeks a Senior Site Reliability Engineer (SRE) to ensure high availability and performance of large-scale cloud infrastructure across AWS and GCP. You will operate distributed data platforms, manage databases such as Cassandra and Elasticsearch, and own incident management lifecycle from detection to RCA.

This onsite role requires strong automation and collaboration with cross-functional teams.

Qualifications

  • Bachelor's or master's degree in Computer Science, Information Systems, or a related field.

Responsibilities

  • Ensure high availability, reliability, and performance of large-scale cloud infrastructure across AWS and GCP environments.
  • Operate and support infrastructure components and distributed data platforms such as Kubernetes, Kafka, Flink, Storm, and Spark.
  • Manage and maintain databases including Cassandra, Elasticsearch, Redis, Postgres, and ArangoDB.
  • Monitor systems, troubleshoot issues, and resolve production incidents across microservices and distributed systems.
  • Collaborate with software engineering teams to debug and resolve complex production problems.
  • Participate in 24x7 on-call rotation supporting multi-cloud production environments.
  • Monitor system metrics, application performance, and infrastructure health using observability tools.
  • Own the incident management lifecycle, including detection, mitigation, RCA, and post-incident reviews.
  • Develop and maintain runbooks, automation, and operational processes to improve reliability and efficiency.
  • Perform capacity planning using system usage and performance data.
  • Drive SRE best practices, operational standards, and continuous improvement initiatives.

Skills

Cloud platforms
Kubernetes
CI/CD
Monitoring
Python

Education

Bachelor's or Master's in CS/IS

Tools

Docker
Terraform
Ansible
Jenkins
GitHub Actions

Job description

Senior Site Reliability Engineer (SRE)

This role has been designed as 'Onsite' with an expectation that you will primarily work from an HPE office.

Who We Are:

Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today’s complex world. Our culture thrives on finding new and better ways to accelerate what’s next. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good. If you are looking to stretch and grow your career our culture will embrace you. Open up opportunities with HPE.

Job Description:
What You Will Do:
  • Ensure high availability, reliability, and performance of large-scale cloud infrastructure across AWS and GCP environments.
  • Operate and support infrastructure components and distributed data platforms such as Kubernetes, Kafka, Flink, Storm, and Spark.
  • Manage and maintain databases including Cassandra, Elasticsearch, Redis, Postgres, and ArangoDB.
  • Monitor systems, troubleshoot issues, and resolve production incidents across microservices and distributed systems.
  • Collaborate closely with software engineering teams to debug and resolve complex production problems.
  • Participate in 24x7 on-call rotation supporting multi-cloud production environments.
  • Monitor system metrics, application performance, and infrastructure health using observability tools.
  • Own the incident management lifecycle, including detection, mitigation, Root Cause Analysis (RCA), and post-incident reviews.
  • Develop and maintain runbooks, automation, and operational processes to improve reliability and efficiency.
  • Perform capacity planning using system usage and performance data.
  • Drive SRE best practices, operational standards, and continuous improvement initiatives.
What You Need to Bring:
  • Bachelor’s or Master’s degree in Computer Science, Information Systems, or a related field.
  • 6-10+ years of experience in DevOps, Site Reliability Engineering, or cloud infrastructure roles.
  • Strong hands-on experience with cloud platforms (AWS or GCP) including services like EC2/GCE, IAM, and object storage (S3/GCS).
  • Experience with containerization and orchestration technologies, especially Docker and Kubernetes.
  • Experience building and managing CI/CD pipelines using tools such as Jenkins, GitHub Actions, or GitLab.
  • Experience with monitoring and observability tools such as Prometheus, CloudWatch, or Stackdriver.
  • Strong understanding of Linux systems administration and configuration management tools like Ansible.
  • Experience managing distributed systems and streaming platforms such as Kafka, Cassandra, Elasticsearch, Spark, Flink, or Storm.
  • Strong automation and scripting skills using Python, Go, Rust, or Shell scripting.
  • Experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
  • Excellent analytical, troubleshooting, and problem-solving skills.
  • Strong communication and collaboration skills with the ability to work with cross-functional teams.
What We Can Offer You:

Health & Wellbeing

We strive to provide our team members and their loved ones with a comprehensive suite of benefits that supports their physical, financial and emotional wellbeing.

Personal & Professional Development

We also invest in your career because the better you are, the better we all are. We have specific programs catered to helping you reach any career goals you have - whether you want to become a knowledge expert in your field or apply your skills to another division.

Unconditional Inclusion

We are unconditionally inclusive in the way we work and celebrate individual uniqueness. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good.

Let's Stay Connected:

Follow @HPECareers on Instagram to see the latest on people, culture and tech at HPE.

#india

#networking

Job:

EngineeringJob Level:

TCP_04

HPE is an Equal Employment Opportunity/ Veterans/Disabled/LGBT employer. We do not discriminate on the basis of race, gender, or any other protected category, and all decisions we make are made on the basis of qualifications, merit, and business need. Our goal is to be one global team that is representative of our customers, in an inclusive environment where we can continue to innovate and grow together.

Hewlett Packard Enterprise is EEO Protected Veteran/ Individual with Disabilities.

HPE will comply with all applicable laws related to employer use of arrest and conviction records, including laws requiring employers to consider for employment qualified applicants with criminal histories.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer, SRE
Senior Site Reliability Engineer, SRE

Hewlett Packard Enterprise • Bengaluru

On-site
INR 3,500,000 - 6,000,000
Health & Wellbeing
Personal & Professional Development
Unconditional Inclusion
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Hewlett Packard Enterprise • Bengaluru Urban

On-site
INR 2,500,000 - 5,200,000
Principal Cloud Developer
Principal Cloud Developer

Hewlett Packard Enterprise Development LP • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Health & Wellbeing
Personal & Professional Development
Unconditional Inclusion
Operations Head
Operations Head

Hewlett Packard Enterprise • Mumbai

On-site
INR 4,500,000 - 6,500,000
Principal Software Engineer
Principal Software Engineer

Hewlett Packard Enterprise Development LP • India

Hybrid
INR 4,000,000 - 6,500,000
Health & Wellbeing
Personal & Professional Development
Unconditional Inclusion
System Software Engineer III (Python Automation, Multi-Features Test (MFT), MPLS (RSVP/LDP), SR[...]
System Software Engineer III (Python Automation, Multi-Features Test (MFT), MPLS (RSVP/LDP), SR[...]

Hewlett Packard Enterprise • Bengaluru

On-site
INR 1,800,000 - 3,600,000
Principal Cloud Developer
Principal Cloud Developer

Hewlett Packard Enterprise • YSR Kadapa

Hybrid
INR 3,500,000 - 5,200,000
Test Automation Engineer
Test Automation Engineer

Hewlett Packard Enterprise India Private Limited • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Health benefits
Development programs
Inclusive culture
Software Systems Engineer-Devops
Software Systems Engineer-Devops

Earn Modes • Bengaluru

On-site
INR 600,000 - 900,000
Managed Services – Azure Local SME - Expert
Managed Services – Azure Local SME - Expert

Hewlett Packard Enterprise • Bengaluru

Hybrid
INR 3,500,000 - 5,500,000
Health & Wellbeing
Career Development Programs
Inclusive Culture