Lead Distributed Systems Engineer - Services Special Projects

Socket.dev

Cupertino (CA)

On-site

USD 250,000 - 350,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple's Special Services Team seeks a Lead Distributed Systems Engineer to design and build massively scalable, low-latency backend services for customer experiences. You will own real-time to analytics workloads, and push forward a multi-tenant platform including AI/ML capabilities.

You will advance distributed-systems best practices from design to production, ensuring high availability, security, and performance across large-scale services.

Qualifications

  • Master's degree and 15+ years of software development experience in scalable distributed systems.
  • Experience in principal or senior staff roles with multi-tenant, high-availability services.
  • Strong concurrency, data structures, and algorithm design knowledge.
  • Proficiency in Java; knowledge of Go or C++ is a plus.
  • Experience with Spring Boot and JVM performance tuning.

Responsibilities

  • Design and build massively scalable, high-availability backend services.
  • Ingest, process, and serve data at scale across workloads from real-time to analytics.
  • Drive evolution of a multi-tenant platform including AI/ML-powered services.

Skills

Distributed systems
Java
Go
C++
Spring Boot
JVM tuning
Async/reactive stacks
gRPC
Kafka
AWS
Kubernetes
Gradle
CI/CD

Education

Master's degree in Computer Science or related field

Tools

Gradle
Jenkins
GitHub Actions
Docker
Kubernetes

Job description

Apple's Special Services Team is seeking a Lead (Principal) Distributed Systems Engineer to design and build massively scalable, highly available services that power experiences for Apple customers both now and in the future.

Description

In this Lead role, you will build and operate high-throughput, low-latency backend services that ingest, process, and serve data at scale across a range of mission-critical workloads — from real-time transactions to analytics and content delivery. You'll also drive the evolution of a multi-tenant platform, including AI/ML-powered services, by shipping new capabilities, scaling what exists, and applying distributed-systems best practices from design through production.

Minimum Qualifications
  • Master's degree in Computer Science or a related field15+ years of professional software development experience building scalable, distributed systems in production, with at least 5 in a Principal or Sr. Staff Level role. Experience building, authoring, and operating large-scale, multi-tiered distributed systems and customer-facing web services: including API design, authentication, authorization, scaling for high availability, concurrency, and reliability.
  • Strong understanding of concurrency and multi-threaded programming, fundamental data structures, and efficient algorithm design
  • Strong proficiency in Java; working knowledge of a second systems language (Go, C++) is a plus.
  • Solid OO analysis and design skills.
  • Strong proficiency in application frameworks (Spring boot)
  • Hands on experience with JVM performance tuning and profiling - GC selection/tuning, JFR, async-profiler, heap/thread-dump analysis for low-latency services.
  • Hands-on with async and reactive JVM stacks: Netty, Project Reactor, RxJava, Vert.x, or Micronaut/Quarkus.
  • Rigorous testing discipline with JUnit 5, Mockito, AssertJ, Testcontainers, and contract testing.
  • Hands on experience with build and dependency management with Gradle
  • Deep understanding of transactional consistency models - ACID semantics, with deep knowledge of tradeoffs between relational and NoSQL database technologies
  • Expertise with synchronous and asynchronous network I/O and RPC frameworks (gRPC)
  • Experience building and maintaining CI/CD pipelines (e.g., Jenkins, GitHub Actions, GitLab CI, or similar) for automated testing, build, and deployment of production services.
  • Experience with AWS or GCP and cloud-native tooling (Docker, Kubernetes) in the context of deploying scalable production grade services.
  • Experience with event streaming and queueing systems (specifically Kafka) and stream processing frameworks and high-throughput, append-only write paths for durable, queryable historical records.
  • Hands on Experience of leveraging data storage (Iceberg, Cassandra) and caching technologies (Redis) in Production services
  • Hands on experience with Serialization/schema tooling: Jackson, Protobuf, Avro, and Schema Registry
  • Experience identifying, triaging, and remediating security vulnerabilities in production services (dependency management, secure code review, threat modeling).
  • Full life-cycle development experience for a consumer product, from concept through deployment
  • Proven history of presenting technical and business concepts to Executive Leadership
Preferred Qualifications
  • Self-motivated, with strong collaboration and communication skills, and experience in a fast-paced, agile environmentExperience with machine learning systems, ML frameworks, libraries and algorithms
  • Familiarity with deployment and optimization of Large scale Production grade AI Services that require GPUs in the path of the transaction.
  • Hands-on experience deploying, serving, and optimizing LLMs or ML models directly in the transaction/request path
  • Experience with security and cryptography (e.g., TLS, X.509 certificates) identity and access management protocols (OAuth2/OIDC/SAML), and secure token/session lifecycle management.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Distributed Systems Engineer - Services Special Projects
Senior Distributed Systems Engineer - Services Special Projects

Socket.dev • Cupertino (CA)

On-site
USD 170,000 - 210,000
Principal Distributed Systems Architect
Principal Distributed Systems Architect

Socket.dev • Cupertino (CA)

On-site
USD 250,000 - 350,000
Core Services Media API Engineering Manager - Platforms Team
Core Services Media API Engineering Manager - Platforms Team

Apple • Cupertino (CA)

On-site
USD 200,000 - 260,000
Senior Software Engineer, Apple Data Platform
Senior Software Engineer, Apple Data Platform

Socket.dev • Cupertino (CA)

On-site
USD 150,000 - 190,000
Software Engineering Technical Lead, Server Energy Services
Software Engineering Technical Lead, Server Energy Services

Socket.dev • San Diego (CA)

On-site
USD 180,000 - 260,000
Principal Data Architect and Manager - Service Special Projects
Principal Data Architect and Manager - Service Special Projects

Socket.dev • Cupertino (CA)

On-site
USD 180,000 - 260,000
Sr. Data Engineer - Services Special Project
Sr. Data Engineer - Services Special Project

Socket.dev • Cupertino (CA)

On-site
USD 180,000 - 240,000
Senior Engineering Manager - Solutions Engineering, Apple Data Platform
Senior Engineering Manager - Solutions Engineering, Apple Data Platform

Socket.dev • Cupertino (CA)

On-site
USD 210,000 - 320,000
Senior Software Engineer - Distributed Systems
Senior Software Engineer - Distributed Systems

Apple Inc. • Cupertino (CA)

On-site
USD 147,400 - 272,100
Comprehensive medical and dental coverage
Employee stock programs
Educational reimbursement
Senior AI Engineer - Services Special Projects
Senior AI Engineer - Services Special Projects

Socket.dev • Cupertino (CA)

On-site
USD 190,000 - 270,000