Site Reliability Engineer, Apple Data Platform / Big Data Platform

Socket.dev

Austin (TX)

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple seeks a Site Reliability Engineer for its Data Platform, focusing on big data engines and catalog/governance layers used to power analytics across the company. You will operate and support the team’s portfolio from ML/AI platform services to multi-cloud infrastructure and collaborate with internal customers daily.

You will be the go-to expert for Spark, Flink, Trino, Notebooks, and catalog services, translating complex issues into clear resolutions while owning reliability roadmaps and

Qualifications

  • Bachelor's Degree in Computer Science, an engineering-related field, or equivalent related experience.
  • 1-4 years in a Site Reliability Engineering, DevOps, or Infrastructure-focused role.
  • Proficient in Python; working knowledge of Golang a plus.
  • Deep understanding of one or more Big Data technologies (Spark, Flink, Airflow, Trino, Notebooks).
  • Experience with Kubernetes and at least one major cloud provider (AWS or GCP).
  • Excellent written and verbal communication skills, with the ability to explain technical issues clearly to non-expert customers.
  • Solid grounding in SRE principles, with prior on-call, production-support, or customer-facing support role experience.

Skills

Python
Golang
Kubernetes
AWS
GCP
Big Data
Spark
Flink
Airflow
Trino
Notebooks
Communication
SRE principles
On-call experience

Education

Bachelor's Degree in Computer Science or equivalent

Tools

Prometheus
Grafana
Splunk
PagerDuty
Glue Catalog

Job description

The Apple Services Engineering team (ASE) is one of the most exciting examples of Apple's long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, and Apple Books — at extensive scale, meeting high expectations to deliver a huge variety of entertainment in over 35 languages to more than 150 countries. Within ASE, the Apple Data Platform SRE team keeps a massive, multi-cloud platform running for thousands of internal engineers building the next generation of data and AI products at Apple. We sit at the intersection of infrastructure, automation, and customer success — running incident response, providing hands‑on support to internal teams, and partnering with developers to make cutting‑edge services like Spark, Flink, Airflow, Trino, Notebooks, and LLM‑based agent platforms reliable at scale.

DESCRIPTION

This is a rare opportunity to build deep expertise across one of the most technically diverse platforms at Apple — while specialising in the big data engines and catalog/governance layers that power analytics and data engineering across the company. As an SRE on Apple Data Platform, you'll operate and support the team's full portfolio, from ML/AI platform services to multi‑cloud infrastructure, and grow into the team's go‑to expert for big data platform services — including Spark, Flink, Airflow, Trino, Notebooks, REST Catalog services (such as Glue Catalog), and data governance. Just as importantly, you'll be a first point of contact for the internal customers who rely on these services daily — someone who can translate a confusing error or a vague support request into a clear diagnosis and a fast resolution. We're looking for a self‑motivated engineer who thrives on ownership — someone who wants a set of services to call their own, the autonomy to drive their reliability roadmap, and the collaborative instinct to keep that work aligned with the team's broader direction. If you love solving hard operational problems, take genuine satisfaction in helping frustrated customers get unblocked, and want a front‑row seat to how Apple's data engineering platform scales, this role offers real room to grow your scope and impact over time.

MINIMUM QUALIFICATIONS
  • Bachelor's Degree in Computer Science, an engineering‑related field, or equivalent related experience.
  • 1-4 years in a Site Reliability Engineering, DevOps, or Infrastructure‑focused role.
  • Proficient in Python; working knowledge of Golang a plus.
  • Deep understanding of one or more Big Data technologies (Spark, Flink, Airflow, Trino, Notebooks).
  • Experience with Kubernetes and at least one major cloud provider (AWS or GCP).
  • Excellent written and verbal communication skills, with the ability to explain technical issues clearly to non‑expert customers.
  • Solid grounding in SRE principles, with prior on‑call, production‑support, or customer‑facing support role experience.
PREFERRED QUALIFICATIONS
  • Experience with REST Catalog services (e.g., Glue Catalog) and data governance frameworks.
  • Prior experience in a customer-facing or technical support role, with a demonstrated passion for customer success.
  • Familiarity with observability tooling: Prometheus, Grafana, Splunk, PagerDuty.
  • Working knowledge of CI/CD pipelines and deployment workflows.
  • Experience with S3 and cloud storage/networking fundamentals.
  • Familiarity with data pipeline orchestration and workflow scheduling patterns.
  • A track record of automating manual operations through scripting or tooling.
  • Intellectual curiosity and a drive to keep learning — for yourself, your team, and the org.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer, Apple Data Platform / Multi-Cloud Infrastructure
Site Reliability Engineer, Apple Data Platform / Multi-Cloud Infrastructure

Socket.dev • Austin (TX)

On-site
USD 140,000 - 220,000
Site Reliability Engineer, Apple Data Platform - AI/ML Platform
Site Reliability Engineer, Apple Data Platform - AI/ML Platform

Apple • Austin (TX)

On-site
USD 120,000 - 180,000
Site Reliability Engineer, Apple Data Platform
Site Reliability Engineer, Apple Data Platform

Socket.dev • Austin (TX)

On-site
USD 150,000 - 230,000
Apple Services Engineering (ASE) Compute - Software Engineering Manager
Apple Services Engineering (ASE) Compute - Software Engineering Manager

Socket.dev • Cupertino (CA)

On-site
USD 190,000 - 240,000
Senior Software Engineer, Apple Data Platform
Senior Software Engineer, Apple Data Platform

Socket.dev • Cupertino (CA)

On-site
USD 150,000 - 190,000
Senior Site Reliability Engineer, Storage SRE / Apple Services Engineering
Senior Site Reliability Engineer, Storage SRE / Apple Services Engineering

Apple Inc. • Cupertino (CA)

On-site
USD 181,000 - 319,000
Comprehensive medical and dental coverage
Retirement benefits
Employee stock programs
+1
Site Reliability Engineer - Kafka
Site Reliability Engineer - Kafka

Socket.dev • Seattle (WA)

On-site
USD 150,000 - 190,000
Senior Software Engineer, Apple Data Platform
Senior Software Engineer, Apple Data Platform

Apple Inc. • Cupertino (CA)

On-site
USD 150,000 - 278,000
Medical and dental coverage
Retirement benefits
Stock programs and RSUs
+1
Software Engineer, Cloud Services ASE
Software Engineer, Cloud Services ASE

Socket.dev • Austin (TX)

On-site
USD 180,000 - 240,000
Service Reliability Engineer (SRE)
Service Reliability Engineer (SRE)

Apple Inc. • Seattle (WA)

On-site
USD 142,000 - 264,000
Medical and Dental coverage
Retirement benefits
Employee stock purchase plan
+2