Site Reliability Engineer/ Expert/ Specialist (Must have strong experience in Windows Server, A[...]

SITA

Delhi

Hybrid

INR 3,500,000 - 7,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flex Week: work from home up to 2 days
Flex Location: up to 30 days travel
Employee Wellbeing program
Professional Development
Competitive Benefits

Job summary

SITA is seeking a Senior Site Reliability/DevOps professional in Delhi to ensure high availability, pro-active incident handling, and robust automation across Azure and on‑prem environments.

You will lead RCAs, design scalable monitoring, and collaborate with cross‑functional teams to improve service readiness and performance. Strong scripting, IaC, and cloud skills are essential for this enterprise role.

Qualifications

  • Bachelor’s degree in computer science, IT, or engineering.
  • 5+ years in IT operations, service management, or infrastructure.
  • Proven experience with high-availability systems on Azure/Windows.
  • RCA and incident management experience.
  • Hands-on with CI/CD pipelines, automation, and IaC.
  • Cross-functional collaboration across Dev/Operations/Engineering.
  • Deployments, risk assessments, and event management.
  • Cloud, containerization, and zero-downtime deployment knowledge.

Responsibilities

  • Build and maintain reliable support systems for high availability.
  • Manage complex operational incidents and root causes.
  • Define event catalogs, alerts, thresholds, and remediation actions.
  • Implement automation for provisioning, monitoring, deployment, self-healing.
  • Collaborate with Product, Engineering, Service Architecture, and Operations.
  • Support customer success with reporting and documentation.
  • Contribute to knowledge management resources.
  • Apply data governance and act as SME for data queries.

Skills

IT Operations
Service Management
Site Reliability Engineer
DevOps
Root Cause Analysis
CI/CD Pipelines
Infrastructure as Code
Cloud Technologies
Azure
Windows Server
Unix/Linux
PowerShell
Kubernetes
Terraform
Observability
Prometheus
Grafana
Dynatrace
CloudWatch

Education

Bachelor’s degree in CS/IT/Engineering

Tools

VMware
Terraform
AKS
Azure Monitor
Dynatrace
Grafana
Prometheus
CloudWatch
Red Hat Linux

Job description

Overview

WELCOME TO SITA. At SITA, we keep airports moving, airlines flying smoothly, and borders open. Our technology and communication innovations power the success of the global air travel industry. You'll find us in 95% of international airports, working closely with over 2,500 transportation and government clients. Each partnership brings unique challenges, and we thrive on delivering fresh solutions and cutting‑edge tech to keep operations running like clockwork. We don't just move the world forward - we're proud to be recognized as a Great Place to Work® by 79% of our employees and certified in most of our growing locations. Here, we feel empowered, supported, and inspired to grow. Are you ready to love your job? The adventure begins right here, with you, at SITA.

Purpose

Ensure high product performance, reliability, and stability by proactively supporting products, resolving root causes of incidents, and implementing improvements that prevent recurrence. This role focuses on event management, automation, service deployment, and operational integration to improve efficiency and collaboration across service operations.

What Will You Do
  • Build and maintain reliable support systems to ensure high availability and strong product performance.
  • Manage complex operational cases, incidents, and root cause analysis to deliver permanent fixes.
  • Define and maintain event catalogs, alerts, thresholds, and remediation actions.
  • Implement automation for provisioning, monitoring, deployment, self‑healing, and recovery.
  • Collaborate with Product, Engineering, Service Architecture, and Operations teams to improve service readiness, availability, and performance.
  • Support customer success initiatives through reporting, documentation, communication materials, and process improvements.
  • Contribute to knowledge management resources such as FAQs, training materials, and operational guidance.
  • Apply data governance standards, monitor data quality, and act as a subject matter expert for data‑related queries.
Qualifications
  • Bachelor’s degree in computer science, Information Technology, Engineering, or a related field.
  • 5+ years of experience in IT operations, service management, or infrastructure management, including roles such as Site Reliability Engineer, Problem Manager, or DevOps Manager.
  • Proven experience managing high‑availability systems and ensuring operational reliability with Azure/Windows Environments.
  • Extensive experience in root cause analysis (RCA), incident management, and developing permanent solutions for recurring service disruptions.
  • Hands‑on experience with CI/CD pipelines, automation, system performance monitoring, and infrastructure as code (IaC).
  • Strong background in collaborating with cross‑functional teams (Development, Operations, Engineering, etc.) to improve operational processes and service delivery.
  • Experience managing deployments, conducting risk assessments, and optimizing event and problem management processes.
  • Familiarity with cloud technologies, containerization, and scalable architectures, including zero‑downtime deployment strategies.
Operating Systems
  • Strong hands on expertise with Windows Server’s (AD, GPO, DNS, DHCP).
  • Solid Problem Management and troubleshooting skills.
  • Working knowledge of Unix/Linux (Red Hat).
  • PowerShell Scripting.
Cloud
  • Strong Experience with Azure and AWS.
  • Strong Knowledge and skills in AKS and on prem Kubernetes.
  • Automation experience, including CI/CD pipelines and exposure to Terraform.
Virtualisation
  • Hands on experience with VMware (VCF, VCD).
Storage
  • Experience managing SAN and NAS.
Backup

Background with Veeam and Veritas backup solutions.

Networking

CCNA level networking knowledge.

Databases

Skills in SQL and MongoDB for restore operations and performance tuning.

Observability & Monitoring
  • Strong experience with enterprise monitoring and observability platforms.
  • Hands‑on experience with Prometheus, Grafana, Dynatrace, Azure Monitor, and AWS CloudWatch.
  • Expertise in infrastructure, application, and service monitoring across on‑premises and cloud environments.
  • Experience designing proactive monitoring, alerting, and capacity management solutions to improve reliability and reduce incidents.
  • Knowledge of log management, metrics, dashboards, distributed tracing, and root cause analysis.
  • Experience defining and monitoring SLIs, SLOs, and SLAs to support Site Reliability Engineering (SRE) practices.
  • Strong focus on operational excellence, platform stability, performance optimisation, and incident reduction through observability‑driven insights.
Core Competencies: (Problem Management)

Conduct thorough problem investigations and root cause analyses to diagnose recurring incidents and service disruptions. Coordinate with Incident Management teams and collaborate with PSOs and Engineering/Product teams to implement permanent solutions. Monitor the effectiveness of problem resolution activities and provide regular reporting to ensure continuous improvement.

What We Offer
  • Flex Week: Work from home up to 2 days/week (depending on your team's needs)
  • Flex Day: Make your workday suit your life and plans.
  • Flex‑Location: Take up to 30 days a year to work from any location in the world.
  • Employee Wellbeing: We have got you covered with our Employee Assistance Program (EAP), for you and your dependents 24/7, 365 days/year. We also offer Champion Health - a personalized platform that supports a range of wellbeing needs.
  • Professional Development: At SITA, we believe growth fuels innovation. Our learning ecosystem offers access to world‑class platforms and programs designed to help you thrive. From LinkedIn Learning, Microsoft's Enterprise Skills Initiative, and Airport Council International - available to all employees-to specialized solutions like Pluralsight for technology upskilling, Harvard Business Publishing for people leadership, Stanford for strategic development and many others, we align learning opportunities with your Development Plan and our business priorities. Your development journey is supported every step of the way.
  • Competitive Benefits: Competitive benefits that make sense with both your local market and employment status.

SITA is an Equal Opportunity Employer. We value a diverse workforce. In support of our Employment Equity Program, we encourage women, aboriginal people, members of visible minorities, and/or persons with disabilities to apply and self‑identify in the application process.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer/ Expert/ Specialist
Site Reliability Engineer/ Expert/ Specialist

SITA • Delhi

On-site
INR 1,800,000 - 2,400,000
Flexible work options
Professional development opportunities
Great Place to Work recognition
Lead Site Reliability Engineer/ Expert (Palo Alto & Versa SD‑WAN Experience)
Lead Site Reliability Engineer/ Expert (Palo Alto & Versa SD‑WAN Experience)

SITA • Delhi

Hybrid
INR 350,000 - 700,000
Associate Service Operations Specialist
Associate Service Operations Specialist

SITA • Delhi

On-site
INR 900,000 - 1,300,000
Flex Week
Flex Day
Flex Location
+3
Associate Infrastructure Engineer
Associate Infrastructure Engineer

SITA • Delhi

Hybrid
INR 1,000,000 - 2,000,000
Flex Week: Work from home up to 2 days/week
Flex Location: Work from any location in the world for up to 30 days a year
Employee Assistance Program for wellbeing
+1
Expert Service Operations
Expert Service Operations

SITA • Delhi

Hybrid
INR 1,000,000 - 1,500,000
Flex Week: Work from home up to 2 days/week
Flex Day: Adjust work hours
Flex-Location: Work from any location for 30 days a year
+3
Associate Service Operations Specialist
Associate Service Operations Specialist

SITA • India

Hybrid
INR 900,000 - 1,200,000
Flex Week: WFH up to 2 days/week
Flex Day
Flex‑Location: up to 30 days/year
+2
Associate Field Engineer
Associate Field Engineer

SITA Group • Bengaluru

On-site
INR 360,000 - 600,000
Flex Week: Work from home up to 2 days
Flex-Location: Up to 30 days/year
Employee Wellbeing
+1
Associate Service Operations Specialist
Associate Service Operations Specialist

SITA Group • Delhi

On-site
INR 600,000 - 900,000
Associate Field Engineer
Associate Field Engineer

SITA • Bengaluru

Hybrid
INR 350,000 - 600,000
Flex Week: Work from home up to 2 days
Flex Day
Flex-Location: up to 30 days travel
+3
Associate Field Engineer
Associate Field Engineer

SITA Group • Mumbai

On-site
INR 400,000 - 600,000
Flex Week
Flex Day
Flex-Location
+3