AI Inference Platform Performance Lead

Apple Inc.

Seattle (WA)

On-site

USD 175,000 - 309,000

Full time

28 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical and dental coverage
Retirement benefits
Discounted products and free services
Tuition reimbursement

Job summary

Apple Inc. in Seattle seeks a senior engineer to build tooling, automation, and analysis capabilities that strengthen our AI inference platform.

You will develop performance benchmarking systems, capacity projection models, and data pipelines that inform infrastructure teams and capacity planners. This cross‑functional role sits at the intersection of AI systems performance, distributed infrastructure, and software engineering, translating complex data into actionable guidance for scaling and

Qualifications

  • BS or MS in Computer Science or related technical field.
  • Strong knowledge of AI/ML inference architecture and the performance characteristics of serving systems.
  • 7+ years of experience in performance and infrastructure engineering in distributed systems.
  • 7 years of coding experience in Python, Go, C++, or other programming languages.
  • Experience with automation engineering, tooling, and data pipelines to support engineering workflows.
  • Strong knowledge of GPU/accelerator architecture as it relates to AI workloads.
  • Practical statistical knowledge applicable to performance analysis and forecasting.
  • Excellent communication skills and ability to turn data into clear guidance for infrastructure teams and capacity planners.

Responsibilities

  • Design and build automations to evaluate AI inference performance across hardware generations and configurations.
  • Develop tooling to surface performance trends, regressions, and insights to infrastructure and planning teams.
  • Build projection and forecasting models to support long-term capacity planning decisions.
  • Analyze performance and utilization data to identify bottlenecks, trends, and optimization opportunities.
  • Partner with AI infrastructure engineers, hardware teams, and capacity planners to deliver critical data and tooling.
  • Create and enhance performance analysis workflows to increase team velocity and data reliability.
  • Continuously improve the accuracy, coverage, and usability of performance measurement and analysis systems.

Skills

Performance engineering
Distributed systems
Python
Go
C++
Automation engineering
Data pipelines
GPU architecture
Statistics
Communication

Education

BS or MS in Computer Science or related technical field

Tools

Nsight
Prometheus
Grafana
Kubernetes
TensorRT-LLM

Job description

Apple Inc. in Seattle seeks a senior engineer to build tooling, automation, and analysis capabilities that strengthen our AI inference platform.

You will develop performance benchmarking systems, capacity projection models, and data pipelines that inform infrastructure teams and capacity planners. This cross‑functional role sits at the intersection of AI systems performance, distributed infrastructure, and software engineering, translating complex data into actionable guidance for scaling and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Inference Platform Performance Engineer
AI Inference Platform Performance Engineer

Socket.dev • Seattle (WA)

On-site
USD 180,000 - 240,000
Senior AI Inference Platform Engineer — Benchmarking
Senior AI Inference Platform Engineer — Benchmarking

Apple • Seattle (WA)

On-site
USD 175,000 - 309,000
Employee stock programs
Employee Stock Purchase Plan
Medical and dental coverage
+4
AI Inference Platform Engineer
AI Inference Platform Engineer

Socket.dev • Seattle (WA)

On-site
USD 180,000 - 240,000
Sr. AI Inference Platform Engineer
Sr. AI Inference Platform Engineer

Apple • Seattle (WA)

On-site
USD 175,000 - 309,000
Employee stock programs
Employee Stock Purchase Plan
Medical and dental coverage
+4
Sr. AI Inference Platform Engineer
Sr. AI Inference Platform Engineer

Apple Inc. • Seattle (WA)

On-site
USD 175,000 - 309,000
Medical and dental coverage
Retirement benefits
Discounted products and free services
+1
Senior AI/ML System Performance Engineer – Architecture
Senior AI/ML System Performance Engineer – Architecture

Socket.dev • Santa Clara (CA)

On-site
USD 180,000 - 240,000
AI Inference Architect & Performance Engineer
AI Inference Architect & Performance Engineer

ElastixAI INC. • Seattle (WA)

Hybrid
USD 180,000 - 240,000
Competitive compensation
Comprehensive medical, dental, and vision coverage
Flexible Time Off (FTO)
+4
AI-Driven Performance Systems Engineer & Agentic Tools
AI-Driven Performance Systems Engineer & Agentic Tools

Apple Inc. • Cupertino (CA)

On-site
USD 129,000 - 225,000
AI Inference Engineer - Performance & API
AI Inference Engineer - Performance & API

Topazlabs • Dallas (TX)

On-site
USD 120,000 - 180,000
Full medical/dental/vision coverage
15 days PTO
5 personal days + holidays
+1
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy