Senior Site Reliability Engineer

Hdfc Bank

Bengaluru

On-site

INR 2,500,000 - 4,000,000

Full time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

HDFC Bank is seeking a Senior Site Reliability Engineer to analyse, troubleshoot, and design vital services and infrastructure with a focus on reliability, scalability, resilience, security, and performance. You will work on cloud-based SaaS environments, automate tasks, and drive observability across containers and backend systems.

The role requires 8–10 years of experience, strong Linux and networking fundamentals, and expertise in Docker/Kubernetes, Terraform, and CI/CD pipelines.

Qualifications

  • 8–10 years of total experience in site reliability or related field.
  • Experience with large-scale infrastructure and cloud-based SaaS environments.
  • Strong Linux, networking, automation, and observability skills.

Responsibilities

  • Help build an SRE culture by sharing best practices and documenting approaches.
  • Automate manual tasks and improve deployment pipelines.
  • Troubleshoot cross-platform issues across OS, networking, and databases in a cloud SaaS setup.
  • Monitor and improve application performance and reliability, and implement fixes.
  • Conduct system analysis and develop improvements for reliability and uptime.
  • Design and ship software to improve observability and efficiency of systems.
  • Maintain deployment orchestration of servers, containers, databases, and backend infra.
  • Develop Run Books and SOPs for recurring Production issues and long-term fixes.
  • Perform incident analysis to prevent future incidents.

Skills

Linux fundamentals
Networking fundamentals
Python
Jenkins
Terraform
Argo CD
AWS
Docker
Kubernetes
Podman
Helm
EKS upgrades
CI/CD
SAST
SCA
Patch management
Monitoring

Education

B Tech in Computer Science

Tools

Terraform
CloudFormation
Ansible
Jenkins
Argo CD
AWS

Job description

Job Title:

Senior Site Reliability Engineer



Job Details:

Business Unit: Tech & Digital


Team: DTIT - Enterprise Factory


Reports to: Lead Site Reliability Engineer


Location: Mumbai, Chennai, Gurgaon & Bangalore


Role Type: Individual Contributor


No of direct reportees: NIL


Travel Required: No


Job Band Range: E4



Job Purpose:

Analysing, troubleshooting, and designing vital services, platforms, and infrastructure while always thinking about reliability, scalability, resilience, security, and performance.



Job Responsibilities:

Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.


Apply automation and software to any tasks or parts of the system which are performed manually.


Able to troubleshoot complicated, cross-platform issues handling OS, Networking, Database in a cloud-based SaaS environment and handle live production incidents.


Monitor application performance, take steps to improve overall application performance and stability, and follow through with implementation.


Conduct system analysis, configuration management, and develop improvements for system software performance, availability, and reliability.


Design, write, ship, and motivate the creation of software and systems to increase observability, product reliability, and organizational efficiency.


Maintain and monitor deployment, orchestration, of servers, docker containers, databases, and general backend infrastructure.


Develop Run Books/Standard Operating Procedure for recurring Production issues, also working on a permanent solve.


Perform Incident Analysis on a regular basis with the intention of preventing and finding a long-term solve for Incidents.



Educational Qualifications:

B Tech in Computer Science or related discipline preferred.



Key Skills:

Experience in monitoring and analyzing infrastructure performance using standard performance monitoring tools.


Demonstrable experience in Containerization-Docker and orchestration (Kubernetes).


Experience with Infrastructure As Code (Terraform, Cloud Formation, Ansible).


Knowledge and proven hands-on experience in large-scale databases and distributed technologies, such as Kafka and Confluent Platform Kafka.


Basic programming and scripting skills.



Experience Required:

Total Yrs of experience: 8-10



Major Stakeholders:

Internal:


Product Manager from Digital Factory


Business Analyst from BTG team


Incident Management team


Development Team



Required Skills


  1. Strong Linux and Networking Fundamentals

  2. Programming concepts - Python is preferred

  3. Jenkins, Terraform, Argo CD, AWS

  4. Docker, Kubernetes, Podman, Helm

  5. Familiarity with EKS Upgrades and pitfalls

  6. CI/CD, SAST, SCA, Patch Management, Monitoring and ing

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Tech & Digital-Lead Site Reliability Engineer
Tech & Digital-Lead Site Reliability Engineer

Hdfc Bank • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Site Reliability Engineer
Site Reliability Engineer

MNR Solutions Pvt. Ltd. • Bengaluru

On-site
INR 900,000 - 1,500,000
Site Reliability Engineer (Virtual Drive - Friday)
Site Reliability Engineer (Virtual Drive - Friday)

Tecnoprism • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Tech S and T-Resilience and Reliability Engineer-Senior-GDSF02
Tech S and T-Resilience and Reliability Engineer-Senior-GDSF02

Ernst & Young Advisory Services Sdn Bhd • Dadri

On-site
INR 3,000,000 - 6,000,000
Senior Site Reliability Engineer - Cloud Infrastructure
Senior Site Reliability Engineer - Cloud Infrastructure

WITS Innovation Lab • Chandigarh

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineer
Site Reliability Engineer

Nielseniq India • Pune District

Hybrid
INR 1,500,000 - 2,100,000
Challenging work content
Professional growth opportunities
Competitive terms of employment
+2
Analyst II, Production Support
Analyst II, Production Support

fis • Pune District

On-site
INR 1,500,000 - 2,300,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Datum Technologies Group • Chennai District

Hybrid
INR 1,000,000 - 1,500,000
Site Reliability Engineer
Site Reliability Engineer

Zorba AI • Chennai District

On-site
INR 4,000,000 - 7,500,000