Senior Site Reliability Engineer, SRE

Hewlett Packard Enterprise

Bengaluru

On-site

INR 3,500,000 - 6,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health & Wellbeing
Personal & Professional Development
Unconditional Inclusion

Job summary

Hewlett Packard Enterprise Development LP in Bengaluru is seeking a Senior Site Reliability Engineer (SRE) for onsite work to help scale our cloud infra across AWS and GCP.

You will own incident response, drive automation with Kubernetes, Docker, Terraform, and CI/CD pipelines, and collaborate with software teams to boost reliability and performance.

Qualifications

  • Bachelor's or Master’s degree in CS/IS.
  • 6–10+ years in DevOps/SRE or cloud infra roles.
  • Experience with AWS or GCP and EC2/GCE.
  • Experience with Kubernetes and Docker.
  • Experience building CI/CD pipelines (Jenkins, GitHub Actions, GitLab).
  • Experience with monitoring/observability tools (Prometheus, CloudWatch, Stackdriver).
  • Strong Linux admin and IaC (Terraform/CloudFormation).
  • Strong scripting (Python/Go/Shell).

Responsibilities

  • Ensure high availability, reliability, and performance of large-scale cloud infrastructure across AWS and GCP.
  • Operate and support Kubernetes, Kafka, Flink, Storm, and Spark.
  • Manage databases such as Cassandra, Elasticsearch, Redis, Postgres, and ArangoDB.
  • Monitor systems, troubleshoot issues, and resolve production incidents across microservices and distributed systems.
  • Collaborate with software engineers to debug and resolve production problems.
  • Participate in 24x7 on-call rotation for multi-cloud environments.
  • Own incident management lifecycle including RCA and post-incident reviews.
  • Develop and maintain runbooks, automation, and operational processes to improve reliability.
  • Perform capacity planning using usage and performance data.
  • Drive SRE best practices and continuous improvement initiatives.

Skills

AWS/GCP
Kubernetes
Docker
CI/CD
Monitoring
Linux admin
Scripting (Python/Go)
IaC (Terraform)
Distributed systems
Collaboration

Education

Bachelor’s or Master’s in CS/IS

Tools

Jenkins
GitHub Actions
GitLab CI
Terraform
CloudFormation
Ansible

Job description

Senior Site Reliability Engineer (SRE) – Onsite, primarily at an HPE office.

Responsibilities
  • Ensure high availability, reliability, and performance of large‑scale cloud infrastructure across AWS and GCP environments.
  • Operate and support infrastructure components and distributed data platforms such as Kubernetes, Kafka, Flink, Storm, and Spark.
  • Manage and maintain databases including Cassandra, Elasticsearch, Redis, Postgres, and ArangoDB.
  • Monitor systems, troubleshoot issues, and resolve production incidents across microservices and distributed systems.
  • Collaborate closely with software engineering teams to debug and resolve complex production problems.
  • Participate in 24x7 on‑call rotation supporting multi‑cloud production environments.
  • Monitor system metrics, application performance, and infrastructure health using observability tools.
  • Own the incident management lifecycle, including detection, mitigation, Root Cause Analysis (RCA), and post‑incident reviews.
  • Develop and maintain runbooks, automation, and operational processes to improve reliability and efficiency.
  • Perform capacity planning using system usage and performance data.
  • Drive SRE best practices, operational standards, and continuous improvement initiatives.
Qualifications
  • Bachelor’s or Master’s degree in Computer Science, Information Systems, or a related field.
  • 6–10+ years of experience in DevOps, Site Reliability Engineering, or cloud infrastructure roles.
  • Strong hands‑on experience with cloud platforms (AWS or GCP), including services like EC2/GCE, IAM, and object storage (S3/GCS).
  • Experience with containerization and orchestration technologies, especially Docker and Kubernetes.
  • Experience building and managing CI/CD pipelines using tools such as Jenkins, GitHub Actions, or GitLab.
  • Experience with monitoring and observability tools such as Prometheus, CloudWatch, or Stackdriver.
  • Strong understanding of Linux systems administration and configuration management tools like Ansible.
  • Experience managing distributed systems and streaming platforms such as Kafka, Cassandra, Elasticsearch, Spark, Flink, or Storm.
  • Strong automation and scripting skills using Python, Go, Rust, or Shell scripting.
  • Experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
  • Excellent analytical, troubleshooting, and problem‑solving skills.
  • Strong communication and collaboration skills with the ability to work with cross‑functional teams.
Benefits
  • Health & Wellbeing – comprehensive suite of benefits supporting physical, financial, and emotional wellbeing.
  • Personal & Professional Development – programs to reach career goals and apply skills across divisions.
  • Unconditional Inclusion – inclusive culture that values diversity.
Equal Employment Opportunity

HPE is an Equal Employment Opportunity / Veterans / Disabled / LGBT employer. We do not discriminate on the basis of race, gender, or any other protected category. All decisions are based on qualifications, merit, and business need. HPE implements E‑Verify and complies with all applicable laws regarding hiring practices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

JobCubby • India

On-site
INR 2,500,000 - 5,500,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Hewlett Packard Enterprise • Bengaluru Urban

On-site
INR 2,500,000 - 5,200,000
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Lyzr AI • Bengaluru

Hybrid
INR 1,000,000 - 2,000,000
Senior SRE Engineer
Senior SRE Engineer

EPAM Systems India Pvt Ltd • Chennai District

On-site
INR 1,800,000 - 3,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

UST • Pune District

On-site
INR 1,800,000 - 3,000,000
Senior Associate Site Reliability Engineer
Senior Associate Site Reliability Engineer

NTT DATA BUSINESS SOLUTIONS • Hyderabad

On-site
INR 1,400,000 - 2,000,000
Senior Software Engineer- Site Reliability
Senior Software Engineer- Site Reliability

S&P Global • Bengaluru

On-site
INR 6,535,000 - 10,271,000
Health & Wellness
Flexible Downtime
Continuous Learning
+3
Senior Software Engineer- Site Reliability
Senior Software Engineer- Site Reliability

S&P Global • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Health care coverage
Generous time off
Access to career resources
+2
Senior Cloud Developer
Senior Cloud Developer

Hewlett Packard Enterprise • YSR Kadapa

On-site
INR 2,500,000 - 4,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobgether • India

On-site
INR 4,200,000 - 6,200,000
Hybrid work in Hyderabad
Health & life insurance