Site Reliability Engineer, Apple Data Platform

Apple Inc.

Austin, Northern (TX, KY)

Hybrid

USD 150,000 - 190,000

Full time

7 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple Inc. is seeking a Site Reliability Engineer for the Data Platform in Austin, TX. You will design, build, and operate scalable data infrastructure supporting services like Apple Music, iCloud, Siri, and Maps.

You’ll work with a cross-functional team to implement reliable deployments, observability, and automated operations across on-premises and cloud environments. Strong coding and collaboration skills are essential.

Qualifications

  • BS/MS in Computer Science or Equivalent.
  • 5+ years of software development or production operations experience in a large-scale environment.
  • Proficiency in authoring and releasing code in Go, Python, or Java using common configuration management and software delivery platforms.
  • Experience operating production applications at scale, including performance testing, HA, DR concepts, capacity planning, and distributed systems on internal and public cloud infrastructure, principally Kubernetes.
  • Understanding of Linux OS, containers and virtualization, networking protocols, and components.
  • Strong ownership and integrity with clear communication and collaboration.
  • Excellent troubleshooting and problem solving using the scientific method.

Responsibilities

  • Solve problems using empirical data, teamwork, and your expertise.
  • Collaborate with partner engineering teams to deliver seamless experiences.
  • Work with open source, vendor licensed, and proprietary tools to improve operations.
  • Apply a consistent incident management process across data platform services and derive SLOs from observability metrics.
  • Balance long-term solutions with business priorities.

Skills

Go
Python
Java
Linux
Troubleshooting

Education

BS/MS in Computer Science or Equivalent

Tools

Kubernetes
AWS
GCP
Ali Cloud

Job description

Site Reliability Engineer, Apple Data Platform

Austin, Texas, United States Software and Services

People at Apple don’t just build products — they craft the kind of experience that have revolutionized entire industries. The diverse collection of our people and their ideas inspire innovation in everything we do. Imagine what you could do here! Join Apple, and help us leave the world better than we found it. Apple Services Engineering (ASE) is responsible for designing and maintaining the systems, platforms, and infrastructure that support Apple's global services, such as Apple Music, iCloud, Siri, Maps, and many more. Our work forms the foundation upon which our world-class software developers build the products our customers love. We are seeking innovative and dedicated Site Reliability Engineers to help us sustain our mission of providing the highest quality experience for our customers. ASE services must scale globally, remain highly available and consistently performant. If you are passionate about designing, engineering, and running systems and infrastructure that will help millions of customers, then this is the place for you!

Description

Apple Services infrastructure is planetary scale. Our Data Platform Site Reliability Engineering team manages the infrastructure and applications on bare-metal and cloud computing platforms to deliver data processing, governance, and storage for many of Apple’s global products and organizations. Our platform teams work with exabytes of data, terabytes of memory, and hundreds of thousands of jobs running millions of executors to support predicable and performant data analytics. Our platform enables key features in Apple Music, TV, Maps, News, and other world class products. Ensuring all of these technologies in geographically distributed data centers work together in harmony presents unique challenges.

Responsibilities
  • You’ll need to solve problems that arise using empirical data, teamwork, and your own unique expertise.
  • Data Platform Services SREs work directly with our partner engineering teams, tightly collaborating with the software developers to deliver seamless experiences for our customers.
  • We run a mix of open source, vendor licensed, and proprietary tools which you will use and have opportunities to improve upon.
  • The cross functional team collaborates to ensure we apply a consistent incident management process across all data platform services and provide user journey based SLOs derived from exhaustive observability metrics, high availability architecture, and automation for deployments.
  • We think critically and strive to balance long-term optimal solutions with the business priorities for each engineering challenge we face. Good ideas are heard and results are rewarded.
Minimum Qualifications
  • BS/MS in Computer Science or Equivalent
  • 5+ years of software development or production operations experience in a large-scale environment
  • Proficiency in authoring and releasing code in Go, Python, or Java using common configuration management and software delivery platforms
  • Experience operating production applications at scale, including well designed performance testing, HA and disaster recovery concepts, capacity planning, and managing distributed systems on internal and public cloud infrastructure, principally Kubernetes
  • Understanding of the Linux Operating System, containers and virtualization, standard networking protocols, and components
  • Strong sense of ownership and integrity demonstrated through clear communication and collaboration
  • Demonstrates excellent troubleshooting and problem solving skills using the scientific method
Preferred Qualifications
  • Proficiency with the architecture, deployment, performance tuning, and troubleshooting of open source data analytics or governance technologies such as Flink, Hive, Hadoop/HDFS, Trino, and/or Druid.
  • Proficiency in managing applications and infra on AWS, GCP and Ali Cloud.
  • The successful candidate is frustrated with toil and has an acute drive to both automate manual operations and evolve them into automatic processes.

Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant

At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.
Learn about accessibility in Apple’s workplace
Learn about reasonable accommodations for job applicants

Apple accepts applications to this posting on an ongoing basis.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer, Apple Data Platform / Big Data Platform
Site Reliability Engineer, Apple Data Platform / Big Data Platform

Apple Inc. • Austin (TX), Northern (KY)

Hybrid
USD 120,000 - 170,000
Site Reliability Engineer, Apple Data Platform / Multi-Cloud Infrastructure
Site Reliability Engineer, Apple Data Platform / Multi-Cloud Infrastructure

Apple Inc. • Austin (TX), Northern (KY)

Hybrid
USD 120,000 - 180,000
Software Engineer, Cloud Services ASE
Software Engineer, Cloud Services ASE

Apple Inc. • Austin (TX), Northern (KY)

Hybrid
USD 140,000 - 190,000
Senior Site Reliability Engineer, Apple Data Platform SRE / Apple Services Engineering
Senior Site Reliability Engineer, Apple Data Platform SRE / Apple Services Engineering

Apple Inc. • Cupertino (CA)

On-site
USD 185,000 - 325,000
Site Reliability Engineer, Apple Data Platform - AI/ML Platform
Site Reliability Engineer, Apple Data Platform - AI/ML Platform

Apple Inc. • Austin (TX), Northern (KY)

Hybrid
USD 140,000 - 180,000
Senior Software Engineer, Apple Data Platform
Senior Software Engineer, Apple Data Platform

Apple Inc. • Cupertino (CA)

On-site
USD 150,000 - 278,000
Medical and dental coverage
Retirement benefits
Stock programs and RSUs
+1
Site Reliability Engineer - ASE Media SRE at Apple Cupertino, CA
Site Reliability Engineer - ASE Media SRE at Apple Cupertino, CA

Apple • Cupertino (CA)

On-site
USD 147,000 - 273,000
Employee stock purchase plan
Restricted stock units (RSUs)
Relocation assistance
+2
Senior Site Reliability Engineer - ASE / iCloud
Senior Site Reliability Engineer - ASE / iCloud

Apple Inc. • San Francisco (CA)

On-site
USD 184,700 - 277,600
Stock programs
Tuition reimbursement
Medical and dental coverage
+2
Senior Site Reliability Engineer - ASE / iCloud
Senior Site Reliability Engineer - ASE / iCloud

Apple Inc. • Seattle (WA)

On-site
USD 175,000 - 308,500
Medical and dental coverage
Retirement benefits
Stock programs (RSUs)
+2
Apple Services Engineering (ASE) Compute - Software Engineering Manager
Apple Services Engineering (ASE) Compute - Software Engineering Manager

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 238,000 - 356,000
Medical and dental coverage
Retirement benefits
Discounted products and services
+1