Site Reliability Engineering Manager

Apple

Bengaluru

On-site

INR 2,000,000 - 2,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple is seeking an experienced Site Reliability Engineering (SRE) Manager in Bengaluru, India, to manage scalable and resilient systems for data pipelines. The role involves technical leadership, driving automation, and ensuring system reliability across both cloud and on-premise environments.

The ideal candidate will have extensive experience in incident management, cloud-native services, and team leadership, helping to maintain operational excellence within engineering teams while optimizing data solutions.

Qualifications

  • 10+ years of experience in Site Reliability Engineering or a related domain.
  • 2+ years of direct people-management experience.
  • Hands-on experience in cloud or hybrid environments.
  • Expertise in cloud-native services and ETL frameworks.

Responsibilities

  • Lead technical leadership and guidance to the SRE team.
  • Drive automation for data platforms and infrastructure.
  • Participate in on-call rotations and resolve incidents.
  • Perform root-cause investigations post-incident.

Skills

Site Reliability Engineering (SRE)
Cloud Computing
Incident Management
Infrastructure as Code (IaC)
Python
Cloud Infrastructure

Education

Bachelor’s degree or equivalent

Tools

Apache Kafka
Apache Spark
Prometheus
Grafana

Job description

Description

We are seeking an experienced Site Reliability Engineering (SRE) Manager to support scalable and resilient distributed systems that power Apple’s data pipelines and analytics platforms. Our Enterprise Data Warehouse landscape caters to a wide variety of real‑time, near real‑time and batch analytical solutions that are integral to business functions such as Sales, Operations, Finance, AppleCare, Marketing and Internet Services. You will work with proprietary and open‑source technologies such as Kafka, Spark, Iceberg, Airflow and others to build these solutions. If you are passionate about addressing infrastructure challenges at scale, both on‑premises and in the cloud, and focused on optimizing scalable solutions by prioritizing ease of use and maintenance, you will discover exciting opportunities in AI & Data Platforms.

As a hands‑on SRE Manager, you’ll lead by example—actively driving operational excellence, contributing to code, and ensuring system reliability. You will be deeply involved in incident response across complex, distributed data platforms designed to support data exploration, analytics and reporting solutions. These platforms operate at the unique intersection of high data volume and hybrid infrastructure, spanning both cloud and on‑premise environments.

Responsibilities
  • Lead by Example: Provide technical leadership and guidance to the SRE team, applying hands‑on skills and continuous learning. Build and mentor a world‑class engineering team that partners closely with platform teams to design scalable, reliable systems, while contributing actively to both platform and application code.
  • Drive Automation for Data Platforms and Infrastructure: Manage Infrastructure as Code (IaC) and develop tooling to enhance engineering productivity. Lead initiatives for cost optimization and operational efficiency at scale.
  • Incident Response and On‑Call Engagement: Actively participate in on‑call rotations and resolve critical production issues. Lead response efforts during major incidents and serve as the primary escalation point for complex problems.
  • Drive Post‑Incident Analysis: Perform root‑cause investigations and ensure follow‑up with actionable post‑mortems and infrastructure hardening initiatives. Implement fixes—in code, infrastructure, or processes—to prevent recurrence.
  • Active Collaboration with Cross‑Functional Teams: Partner closely with engineering teams to troubleshoot issues, deploy fixes, and enhance system reliability. Champion operational excellence through direct technical contributions.
  • Establish Production Readiness Standards: Take ownership of application security, disaster recovery and application documentation to reflect the latest system architecture and configurations.
Minimum Qualifications
  • 10+ years of experience in Site Reliability Engineering (SRE) or a related domain.
  • 2+ years of direct people‑management experience, including leading, hiring, developing, and building engineering teams.
  • Hands‑on experience supporting and maintaining applications in cloud or hybrid environments.
  • Expertise in cloud‑native services, including ETL frameworks (e.g., Apache Spark, Flink) and messaging systems (e.g., Kafka).
  • Strong knowledge of cloud infrastructure & services (e.g., AWS, GCP, Kubernetes).
  • Experience with observability tools (e.g., Prometheus, Grafana, CloudWatch).
  • Programming experience in Python, Java, or Scala.
  • Proven ability to lead incident response, perform root‑cause analysis, and drive system reliability improvements.
  • Bachelor’s degree or equivalent.
Preferred Qualifications
  • Hands‑on experience supporting enterprise data systems on distributed architectures.
  • Exposure to data visualization tools such as Tableau, Business Objects, or ThoughtSpot, with experience supporting and troubleshooting related issues.
  • Experience with modern & distributed databases such as Snowflake, Cassandra, SingleStore, or SAP HANA.
  • Experience using generative AI or automation tools for issue detection, alerting, or remediation.
  • Solid understanding of system design, data structures, and incident management best practices.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE Engineering Manager
SRE Engineering Manager

Apple • Hyderabad

On-site
INR 2,000,000 - 3,000,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

Apple • Hyderabad

On-site
INR 2,500,000 - 3,500,000
Site Reliability Engineering Manager, Apple Data Platform
Site Reliability Engineering Manager, Apple Data Platform

Apple Inc. • Bengaluru

On-site
INR 4,000,000 - 8,000,000
Site Reliability Engineer - Insights
Site Reliability Engineer - Insights

Apple • Bengaluru

On-site
INR 1,200,000 - 2,100,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Synechron • Bengaluru, Hyderabad

Hybrid
INR 4,200,000 - 6,300,000
Site Reliability Engineer (ETS)
Site Reliability Engineer (ETS)

Apple • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Welcoming and diverse work environment
Focus on accessibility and inclusion
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
Service Reliability Engineer- Apple Data Platforms
Service Reliability Engineer- Apple Data Platforms

Apple • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Full Stack Software Engineer - Manufacturing Systems & Infrastructure
Full Stack Software Engineer - Manufacturing Systems & Infrastructure

Apple • Bengaluru

On-site
INR 1,800,000 - 2,500,000