Site Reliability Engineer

Deutsche Bank

Bengaluru

On-site

INR 1,500,000 - 2,700,000

Full time

45 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Leave policy
Parental leave
Childcare reimbursement
Certifications sponsorship
Employee assistance program
Health insurance
Life insurance
Health screening

Job summary

Deutsche Bank Bangalore, India seeks a Site Reliability Engineer to join a role combining production support, SRE, and DevOps. The candidate should have strong GCP experience with Kubernetes (GKE), automation via CI/CD using GitHub Actions, and expertise in Helm charts and service mesh technologies.

You will implement monitoring, incident response, and performance optimization across the stack, collaborating with development teams and maintaining detailed documentation.

Qualifications

  • Proficiency in Terraform (must)
  • Experience with Google Cloud Platform (GCP) and Kubernetes (GKE)
  • Knowledge of security management including GCP Secret Manager
  • Strong Kubernetes administration and operations
  • Experience with CI/CD tools
  • GitHub Actions – CI/CD experience is must
  • Experience with Docker/Kubernetes deployments
  • Experience developing Helm charts
  • Exposure to monitoring tools like Prometheus and Grafana
  • Knowledge of Istio or Anthos Service Mesh
  • Experience in Java, Python, Go, or Bash

Responsibilities

  • Ensure reliability, availability, and performance of production systems with monitoring and incident response best practices
  • Maintain end-to-end application and infrastructure view and support tooling
  • Develop and maintain automation for deployment, scaling, and operations
  • Serve as primary responder to outages, with rapid resolution and post-mortems
  • Design and implement robust monitoring and alerting systems
  • Identify and resolve performance bottlenecks across stack
  • Collaborate with development teams to ensure reliable design of features
  • Document systems, processes, and procedures for knowledge sharing
  • Continuously improve infrastructure and tools for efficiency
  • Design, implement, and manage CI/CD pipelines using GitHub Actions
  • Deploy and operate applications on Google Kubernetes Engine (GKE)
  • Develop and maintain Helm charts for deployments
  • Manage Kubernetes including node management and autoscaling
  • Configure service networking (gateways, virtual services, service mesh)

Skills

Terraform
GCP
Kubernetes
CI/CD
GitHub Actions
Docker
Helm
Istio / Anthos Service Mesh
Java
Python
Go

Tools

GKE
Prometheus
Grafana
Anthos Service Mesh
Helm

Job description

Location: Bangalore, India

Role Description

We are looking for Site Reliability Engineer candidate with below requirement. This role is combination of Production support + SRE + Devops. So majorly looking for GCP experience and Kubernetes to support the design, deployment, automation, and operational excellence of enterprise-grade cloud applications.

Position Overview

Job Title: Site Reliability Engineer
Corporate Title: Assistant Vice President
Location: Bangalore, India

We are looking for Site Reliability Engineer candidate with below requirement. This role is combination of Production support + SRE + Devops. So majorly looking for GCP experience and Kubernetes to support the design, deployment, automation, and operational excellence of enterprise-grade cloud applications.

What We’ll Offer You

As part of our flexible scheme, here are just some of the benefits that you’ll enjoy,

  • Best in class leave policy.
  • Gender neutral parental leaves
  • 100% reimbursement under childcare assistance benefit (gender neutral)
  • Sponsorship for Industry relevant certifications and education
  • Employee Assistance Program for you and your family members
  • Comprehensive Hospitalization Insurance for you and your dependents
  • Accident and Term life Insurance
  • Complementary Health screening for 35 yrs. and above
Your Key Responsibilities
  • System Reliability: Ensure the reliability, availability, and performance of production systems by implementing best practices in monitoring, alerting, and incident response.
  • System Maintenance: Understand thoroughly the end-to-end application support process and escalation procedures, become fully conversant with all support tools. Maintain an end-to-end view of the application and infrastructure landscape.
  • Automation: Develop and maintain automation tools and scripts to streamline deployment, scaling, and operational tasks.
  • Incident Management: Act as a primary responder to system outages and incidents, ensuring rapid resolution and thorough post-mortem analysis to prevent recurrence.
  • Monitoring & Alerting: Design and implement robust monitoring and alerting systems to proactively identify and address potential issues.
  • Performance Optimization: Identify and resolve performance bottlenecks across the stack, from application code to infrastructure.
  • Collaboration: Work closely with development teams and other stakeholders to ensure that new features and services are designed with reliability and scalability in mind.
  • Documentation: Maintain comprehensive documentation of systems, processes, and procedures to ensure knowledge sharing and continuity.
  • Continuous Improvement: Continuously evaluate and improve our infrastructure, tools, and processes to enhance system reliability and operational efficiency.
  • Design, implement, and manage CI/CD pipelines using GitHub Actions.
  • Deploy and operate applications on Google Kubernetes Engine (GKE).
  • Develop and maintain Helm charts for complex application deployments.
  • Manage Kubernetes infrastructure including node management, auto-scaling, configuration management, and secrets management.
  • Configure and support service networking components such as gateways, virtual services, and service mesh technologies (Anthos Service Mesh preferred).
Your Skills And Experience
  • Proficiency in Infrastructure as Code - Terraform (must)
  • Proficiency in cloud platforms such Google Cloud (preferred), Openshift Cloud
  • Usage of enterprise Security Management solutions including GCP Secret Manager.
  • Expertise in Kubernetes (GKE) administration and operations
  • Experience in CI/CD tools
  • GitHub Actions – CI/CD experience is must
  • Experience with Docker/Kubernetes (creating images, deployments)
  • Experience into developing Helm Charts ( templates, hooks, packaging)
  • Exposure to delivering good quality code within enterprise scale development
  • Working knowledge of environment monitoring tools such as GCO, Prometheus, Grafana
  • Strong experience in software development processes, models, lifecycles and methodologies.
  • Expert hands-on experience with service-mesh technology such as Istio or Anthos Service Mesh
  • Experience in software development and scripting in at least one language (Java, JavaScript, Python, Go, Bash)

Proven ability to leverage AI tools to enhance productivity, optimise workflows to solve business problems, while applying critical judgment to ensure responsible and ethical use of data and AI outputs.

How We’ll Support You
  • Training and development to help you excel in your career.
  • Coaching and support from experts in your team.
  • A culture of continuous learning to aid progression.
  • A range of flexible benefits that you can tailor to suit your needs.
About Us And Our Teams

https://www.db.com/company/company.html

We strive for a culture in which we are empowered to excel together every day. This includes acting responsibly, thinking commercially, taking initiative and working collaboratively.

Together we share and celebrate the successes of our people. Together we are Deutsche Bank Group.

We welcome applications from all people and promote a positive, fair and inclusive work environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Talent'd HR Solutions • Pune District

On-site
INR 800,000 - 1,200,000
Industry leading leave policies
Gender neutral parental leave
Full reimbursement through childcare assistance benefits
+5
Site Reliability Engineer
Site Reliability Engineer

Arch Systems • Hyderabad

On-site
INR 2,800,000 - 4,200,000
Site Reliability Engineer- GCP
Site Reliability Engineer- GCP

Aziro • Hyderabad

Hybrid
INR 1,400,000 - 2,100,000
Software Engineer-DevOps
Software Engineer-DevOps

SMC Squared India • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hilabs • Bengaluru

On-site
INR 2,200,000 - 3,500,000
DevOps Lead, AVP
DevOps Lead, AVP

Deutsche Bank • Pune District

On-site
INR 2,500,000 - 5,000,000
Leave policy and parental leave
Certifications sponsorship
Employee Assistance Program
+3
Engineer, AS
Engineer, AS

Deutsche Bank • Pune District

On-site
INR 900,000 - 1,500,000
Leave policy
Parental leave
Childcare assistance reimbursement
+5
Senior Data Site Reliability Engineer | GCP is mandatory
Senior Data Site Reliability Engineer | GCP is mandatory

Anlage Infotech • Chennai District

On-site
INR 3,500,000 - 6,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MNC Group • Kamrup Metropolitan

On-site
INR 4,000,000 - 6,000,000
Engineer, AS
Engineer, AS

Deutsche Bank • Maharashtra

On-site
INR 1,200,000 - 1,800,000
Leave policy
Parental leaves
Childcare assistance
+5