Senior Site Reliability Engineer

Embarkgcc Services

Bengaluru

On-site

INR 1,200,000 - 1,800,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Embarkgcc Services in Bengaluru is seeking a Site Reliability Engineer / Platform SRE to manage Azure-based cloud infrastructure, Kubernetes deployments and production support.

You will troubleshoot real-world production issues, improve platform reliability, monitor applications and collaborate with Platform Engineering and DevOps teams; this role emphasizes operations and observability with tools like Datadog or Dynatrace, and involves CI/CD improvements and basic security tasks.

Qualifications

  • Hands-on experience with Microsoft Azure.
  • Strong practical experience with Kubernetes.
  • Good understanding of SRE / Platform Engineering concepts.
  • Strong production troubleshooting and RCA experience.
  • Experience with CI/CD and deployment processes.
  • Good understanding of monitoring, logging and observability.
  • Exposure to Datadog, Dynatrace or similar observability tools.
  • Experience with Apache Flink is preferred.
  • Understanding of ETL/ELT and data pipelines.
  • Experience working

Responsibilities

  • Manage and support Azure-based cloud infrastructure and services.
  • Work with Kubernetes for deployments, troubleshooting, scaling and production support.
  • Monitor production environments and investigate application/infrastructure issues.
  • Analyze logs, metrics and alerts to identify issues and perform root-cause analysis (RCA).
  • Work closely with Platform Engineering and DevOps teams to maintain reliable production environments.
  • Support and improve CI/CD pipelines and deployment processes.
  • Work with Apache Flink and data-processing environments where required.
  • Support ETL/ELT pipelines and data-platform workflows.
  • Contribute to monitoring and observability using tools such as Datadog, Dynatrace or similar platforms.
  • Participate in incident management and help improve system reliability.
  • Support vulnerability identification, remediation and other basic security-related activities.
  • Assist with API performance/load testing and identifying performance bottlenecks.
  • Use automation and AI-assisted tools to improve operational efficiency.

Skills

Microsoft Azure
Kubernetes
SRE concepts
RCA
CI/CD
Observability
Datadog
Dynatrace
Apache Flink
ETL/ELT
Data pipelines
Incident management
Automation

Job description

About the Role

We are looking for a Site Reliability Engineer / Platform SRE who enjoys working with cloud infrastructure, Kubernetes and production environments.

The ideal candidate should be comfortable troubleshooting real-world production issues, improving platform reliability, monitoring applications and working closely with Platform and DevOps teams.


This is not a coding-heavy role. We are looking for someone who can understand existing code and scripts, troubleshoot issues and make basic changes when required. AI-assisted tools can be used for more complex coding requirements.


Key Responsibilities
  • Manage and support Azure-based cloud infrastructure and services.
  • Work with Kubernetes for deployments, troubleshooting, scaling and production support.
  • Monitor production environments and investigate application/infrastructure issues.
  • Analyze logs, metrics and alerts to identify issues and perform root-cause analysis (RCA).
  • Work closely with Platform Engineering and DevOps teams to maintain reliable production environments.
  • Support and improve CI/CD pipelines and deployment processes.
  • Work with Apache Flink and data-processing environments where required.
  • Support ETL/ELT pipelines and data-platform workflows.
  • Contribute to monitoring and observability using tools such as Datadog, Dynatrace or similar platforms.
  • Participate in incident management and help improve system reliability.
  • Support vulnerability identification, remediation and other basic security-related activities.
  • Assist with API performance/load testing and identifying performance bottlenecks.
  • Use automation and AI-assisted tools to improve operational efficiency.

Required Skills :
  • Hands-on experience with Microsoft Azure.
  • Strong practical experience with Kubernetes.
  • Good understanding of SRE / Platform Engineering concepts.
  • Strong production troubleshooting and RCA experience.
  • Experience with CI/CD and deployment processes.
  • Good understanding of monitoring, logging and observability.
  • Exposure to Datadog, Dynatrace or similar observability tools.
  • Experience with Apache Flink is preferred.
  • Understanding of ETL/ELT and data pipelines.
  • Experience working 
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Zorba AI • Chennai District

On-site
INR 1,200,000 - 2,400,000
Site Reliability Engineer
Site Reliability Engineer

Metlife • Hyderabad

Hybrid
INR 1,500,000 - 2,600,000
Senior Site Reliability Engineer (Azure) - S
Senior Site Reliability Engineer (Azure) - S

Tata Consultancy Services • Kolkata District, Chennai District, Bengaluru

On-site
INR 2,400,000 - 4,200,000
Site Reliability Engineer
Site Reliability Engineer

PwC India • Bengaluru

On-site
INR 1,500,000 - 2,500,000
SRE - AWS, GCP & Azure
SRE - AWS, GCP & Azure

PibyThree • Thane

On-site
INR 1,200,000 - 1,500,000
SRE - AWS, GCP & Azure
SRE - AWS, GCP & Azure

PibyThree • Navi Mumbai

On-site
INR 1,200,000 - 1,800,000
DevOps Engineer/Site Reliability Engineer
DevOps Engineer/Site Reliability Engineer

Thompsons HR Consulting Pvt Ltd • Pune District

On-site
INR 900,000 - 1,300,000
Site Reliability Engineer
Site Reliability Engineer

Ascendion • Chennai District

On-site
INR 2,000,000 - 4,200,000
Site Reliability Engineer (SRE) / DevOps Engineer
Site Reliability Engineer (SRE) / DevOps Engineer

New Era Technology • Gurugram District

On-site
INR 1,500,000 - 2,100,000
Assistant Manager - Azure Site Reliability Engineer
Assistant Manager - Azure Site Reliability Engineer

Promaynov Advisory Services Pvt. Ltd • Bengaluru

On-site
INR 1,400,000 - 2,100,000