Sr DevOps Engineer

METRO France

Wagholi

Hybrid

INR 2,400,000 - 3,600,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

METRO France is seeking a Senior Site Reliability Engineer to design and maintain scalable cloud-native systems in a hybrid setup at MGSC Pune. The role focuses on GCP, Kubernetes, IaC, observability, and automation to ensure reliability and performance across global services.

The ideal candidate will bring hands-on experience with Terraform, Helm, Kustomize, Datadog, and strong scripting abilities, plus excellent English communication skills for collaboration with global teams.

Qualifications

  • 7+ years of experience in Site Reliability Engineering.
  • Experience designing elastic, resilient cloud systems.
  • Strong understanding of GCP and Kubernetes.
  • Proficiency with infrastructure as code and configuration management tools.
  • Experience with monitoring and observability tools.
  • Scripting skills in Bash and automation.
  • Experience with CI/CD pipelines, especially GitHub Actions.
  • Networking fundamentals and troubleshooting.
  • Strong English communication.
  • Ability to work independently in a fast-paced environment.

Responsibilities

  • Ensure the stability and reliability of cloud-native applications deployed on GCP.
  • Define, implement, and monitor SLOs, SLAs, and SLIs.
  • Automate infrastructure provisioning using Terraform and manage Kubernetes configurations with Kustomize and Helm.
  • Develop and maintain monitoring and alerting systems using Datadog and GCP native tools.
  • Conduct incident analysis and postmortems to drive continuous improvement.
  • Collaborate with development teams to integrate reliability practices into CI/CD pipelines using GitHub Actions.
  • Manage and troubleshoot database systems, particularly PostgreSQL and Cassandra.
  • Apply networking knowledge and Linux system administration skills to troubleshoot and optimize system connectivity and performance.

Skills

Site Reliability
GCP
Kubernetes
Terraform
Helm
Kustomize
Datadog
GitHub Actions
Bash scripting
Networking

Education

Bachelor’s or Master’s degree in Computer Science or related field

Tools

Docker
Prometheus
Datadog
GCP Monitoring

Job description

Job Description:
Company Description

Company Description

Metro Global Solution Center (MGSC)is internal solution partner for METRO, a €31.6 Billion international wholesaler with operations in more than 30 countries. The store network comprises a total of 623 stores in 21 countries, of which 522 offer out-of-store delivery (OOS), and 94 dedicated depots. In 12 countries, METRO runs only the delivery business by its delivery companies (Food Service Distribution, FSD).

HoReCa and Traders are core customer groups of METRO. The HoReCa section includes hotels, restaurants, catering companies as well as bars, cafés and canteen operators. The Traders section includes small grocery stores and kiosks. The majority of all customer groups are small and medium-sized enterprises as well as sole traders. METRO helps them manage their business challenges more effectively.

MGSC, location wise is present in Pune (India), Düsseldorf (Germany) and Szczecin (Poland). We provide HR, Finance, IT, Strategy, Branding & Business operations support to 31 countries, speak 24+ languages and process over 18,000 transactions a day. We are setting tomorrow’s standards for customer focus, digital solutions, and sustainable business models. For over 10 years, we have been providing services and solutions from our two locations in Pune and Szczecin. This has allowed us to gain extensive experience in how we can best serve our internal customers with high quality and passion. We believe that we can add value, drive efficiency, and satisfy our customers.

Job Description

Job Description

Role Overview

We are seeking a Senior Site Reliability Engineer with strong experience in building and maintaining scalable, resilient systems. The ideal candidate will have hands‑on expertise in cloud-native technologies, infrastructure as code, observability, and automation, with a focus on Google Cloud Platform (GCP).

Key Responsibilities
  • Ensure the stability and reliability of cloud-native applications deployed on GCP,containerized with Docker and orchestrated via Kubernetes.
  • Define, implement, and monitor SLOs, SLAs, and SLIs to measure system performance and user experience.
  • Automate infrastructure provisioning using Terraform and manage Kubernetesconfigurations with Kustomize and Helm.
  • Develop and maintain monitoring and alerting systems using Datadog and GCP-nativetools.
  • Conduct incident analysis and postmortems to drive continuous improvement.
  • Collaborate with development teams to integrate reliability practices into CI/CDpipelines using GitHub Actions.
  • Manage and troubleshoot database systems, particularly PostgreSQL and Cassandra.
  • Apply networking knowledge and Linux system administration skills to troubleshoot andoptimize system connectivity and performance.
Work Experience & Skills
  • 7+ years of experience in Site Reliability Engineering.
  • Proven experience designing and operating elastic, resilient systems in cloud environments.
  • Strong understanding of GCP, Kubernetes, and container orchestration.
  • Proficiency in infrastructure as code and configuration management tools (Terraform,
  • Helm, Kustomize).
  • Experience with monitoring and observability tools (Datadog, GCP Monitoring).
  • Solid scripting skills in bash and familiarity with automation frameworks.
  • Experience with CI/CD pipelines, especially using GitHub Actions.
  • Familiarity with networking fundamentals and troubleshooting.
  • Strong coding skills and ability to develop reliability-focused tooling.
  • Excellent communication skills in English (written and spoken)
Other Requirements
  • Strong problem‑solving skills and a process‑oriented mindset.
  • Ability to work independently and collaboratively in a fast‑paced environment.
  • Passion for clean code, automation, and continuous improvement.
Nice‑to‑Have
  • Familiarity with monitoring tools (e.g., DataDog, Prometheus, GCP Monitoring).
  • Experience working in Agile/Scrum teams.
Qualifications

Education

  • Bachelor’s or Master’s degree in Computer Science, Software Engineering, orequivalentpracticalexperience.
Requirements

Work Model: Hybrid

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Devops Engineer
Sr. Devops Engineer

METRO/MAKRO • Pune District

On-site
INR 1,800,000 - 3,200,000
Site Reliability Engineer (SRE) – GCP Platform
Site Reliability Engineer (SRE) – GCP Platform

ITC Infotech • Bengaluru

On-site
INR 900,000 - 1,300,000
Senior Engineer DevOps [T500-28630]
Senior Engineer DevOps [T500-28630]

Albertsons Companies India • Bengaluru

On-site
INR 4,000,000 - 6,000,000
DevOps Engineer
DevOps Engineer

Sonata Software • Hyderabad

Hybrid
INR 1,200,000 - 1,800,000
DevOps & Site Reliability Engineer (GCP)
DevOps & Site Reliability Engineer (GCP)

Qureos • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Staff Engineer DevOps [T500-28639]
Staff Engineer DevOps [T500-28639]

Albertsons Companies India • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Senior SecDevOps
Senior SecDevOps

Moneymul Technologies Pvt Ltd • Dadri, Delhi

On-site
INR 1,400,000 - 2,400,000
Senior Data Site Reliability Engineer | GCP is mandatory
Senior Data Site Reliability Engineer | GCP is mandatory

Anlage Infotech • Chennai District

On-site
INR 3,500,000 - 6,000,000
Site Reliability Engineer- GCP
Site Reliability Engineer- GCP

Aziro • Hyderabad

Hybrid
INR 1,400,000 - 2,100,000
Lead SRE / Sr. DevOps Engineer
Lead SRE / Sr. DevOps Engineer

V2 Solutions • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Medical Insurance
OPD Wallet
Parental cover
+4