Senior Site Reliability Engineer

Lex

Singapore

On-site

SGD 180,000 - 240,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Apple's Services Engineering Cloud Services SRE team in Singapore is seeking a strong SRE to help scale the core platform behind Apple internet services, affecting billions of users.

The team builds automation, instrumentation and tools to scale infrastructure, responds to alerts and incidents, and focuses on reliability improvements in hardware, networking, OS and software. The role requires expertise in Kubernetes, distributed systems, and backend services.

Qualifications

  • Bachelor’s or Master’s in Computer Science, Computer Engineering, or equivalent experience.
  • Proficiency in Go, Rust, Python or Swift.
  • Proven backend internet services software development experience.
  • Knowledge of CI/CD, testing methods, TDD and Agile.
  • Understanding of DNS, DHCP, virtualization and monitoring.
  • Experience operating large distributed systems.

Responsibilities

  • Build automation, instrumentation and tools to scale the systems reliably.
  • Respond to alerts and incidents to maintain platform reliability.
  • Improve reliability and efficiency of the services at scale.

Skills

Kubernetes ecosystem experience
UI frameworks (React/Angular)
Large-scale server provisioning

Education

Bachelor’s or Master’s in CS/CE

Tools

OpenStack Ironic
MAAS
Netbox
Tinkerbell
Puppet/Chef/Ansible

Job description

Summary

Become a Site Reliability Engineer in Apple’s Cloud Service Infrastructure team, part of Apple’s Services Engineering organization, and help scale the cloud that underpins services for billions of Apple users.

We are building and supporting new and existing infrastructure to support the hyperscaling of Apple Silicon systems in the datacenter. This allows Apple to provide best in class privacy and power efficiency for AI users (as part of our Private Cloud Compute service) and for all Apple Services. We are at the cutting edge of Apple’s cloud hardware and software infrastructure, moving the dial so that Apple can provide new and exciting services to our end users. We help Apple surprise and delight our users.

Description

The Apple Services Engineering Cloud Services SRE organization is looking for a strong, enthusiastic SRE to join our team in Singapore. This person will have a tremendous amount of individual responsibility and influence over the direction the core platform of many critical Apple internet services takes for years to come. You are someone with ideas and real passion for software delivered as a service to improve reuse, efficiency, and simplicity. This engineer’s work will impact billions of users and be essential to the success of some of the most visible current and future Apple features.

We are domain experts in fleet management, systems, and software engineering. We build automation, instrumentation and tools to scale the systems reliably. We respond to alerts and incidents which may pose a risk to the reliability of the platform and we learn from them to improve the future performance of the services. The team’s focus is on infrastructure capabilities and processes, improving the reliability and efficiency of the systems, at scale.

We have a range of expertise in the team across hardware, networking, distributed systems, reliability, processes, operating systems, software development. We need people who can bring their own expertise to bear and are happy to teach and learn as we grow the service.

Desired Skills
  • Experience with large scale server provisioning and maintenance (OpenStack Ironic, Metal3, MAAS, xCat, Netbox, Tinkerbell)
  • Experience with development within Kubernetes ecosystem, including operator framework, controllers and CRDs
  • Experience with UI frameworks such as React or Angular
Some Exposure To The Following
  • Hardware bootstrap and associated security (PXE, BIOS, TPM, secure boot, trusted computing)
  • Structured or unstructured storage and caching
  • Automating operations processes via services and tools
  • Configuration management and fleet orchestration via Puppet, Chef, Ansible, or others
  • Cloud Services (AWS S3/EC2/CloudFront or equivalent)
Minimum Qualifications
  • Bachelor’s or Master’s in Computer Science, Computer Engineering, or equivalent experience.
  • Strong emphasis on SRE as an engineering subject area, with proficiency in at least one of the following languages (Go, Rust, Python, Swift)
  • A successful track record and proven experience as a backend internet services software developer
  • Knowledge of the software development lifecycle, including continuous integration, testing methodologies, TDD and agile development methodologies
  • Understanding of foundational internet infrastructure services including DNS, DHCP, virtualization and monitoring
  • Experience operating critical, large-scale distributed systems spanning hardware, operating systems, and software
  • Understanding of SRE principles, including observability, alerting, error budgets, fault analysis, and other common reliability engineering concepts, with a keen eye for opportunities to eliminate toil by code and process improvements
Preferred Qualifications
  • Hardware bootstrap and associated security (PXE, BIOS, TPM, secure boot, trusted computing)
  • Structured or unstructured storage and caching
  • Automating operations processes via services and tools
  • Configuration management and fleet orchestration via Puppet, Chef, Ansible, or others
  • Cloud Services (AWS S3/EC2/CloudFront or equivalent)

Apple is an equal opportunity employer that is committed to inclusion and diversity. Apple provides reasonable accommodations to applicants with disabilities and in accordance with local requirements. Apple is a drug-free workplace.

At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.

Learn about accessibility in Apple’s workplace

Role Number: 200674507-3278

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Apple Inc. • Singapore

On-site
SGD 120,000 - 210,000
Site Reliability Engineer, Enterprise Technology Services
Site Reliability Engineer, Enterprise Technology Services

Apple • Singapore

On-site
SGD 120,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Apple • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer, Emerging Technology
Site Reliability Engineer, Emerging Technology

Apple • Singapore

On-site
SGD 60,000 - 90,000
Site Reliability Engineer, Emerging Technology
Site Reliability Engineer, Emerging Technology

Apple Inc. • Singapore

On-site
SGD 90,000 - 150,000
Senior Cloud SRE Engineer – Scale & Reliability
Senior Cloud SRE Engineer – Scale & Reliability

Apple Inc. • Singapore

On-site
SGD 120,000 - 210,000
Senior Software Engineer, Cloud Services Engineering
Senior Software Engineer, Cloud Services Engineering

Apple Inc. • Singapore

On-site
SGD 180,000 - 240,000
Senior Cloud SRE: Scale Infra for Billion-User Services
Senior Cloud SRE: Scale Infra for Billion-User Services

Lex • Singapore

On-site
SGD 180,000 - 240,000
Emerging Tech SRE: Build Reliable, Automated Cloud Ops
Emerging Tech SRE: Build Reliable, Automated Cloud Ops

Apple Inc. • Singapore

On-site
SGD 90,000 - 150,000
Senior SRE & DevOps Engineer — Cloud, Kubernetes, ML
Senior SRE & DevOps Engineer — Cloud, Kubernetes, ML

Apple • Singapore

On-site
SGD 120,000 - 180,000