Senior AI Inference Platform Engineer — Benchmarking

Apple

Seattle (WA)

On-site

USD 175,000 - 309,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Employee stock programs
Employee Stock Purchase Plan
Medical and dental coverage
Retirement benefits
Discounted products
Tuition reimbursement
Discretionary bonuses

Job summary

Apple seeks an engineer to build tooling, automation, and analysis capabilities for the AI inference platform, focusing on benchmarking, capacity projection models, and data pipelines that inform infrastructure teams and capacity planners.

You will design real-time performance dashboards, collaborate with AI infrastructure engineers, hardware teams, and capacity planners to forecast needs months in advance, and help accelerate data-driven decisions for scaling AI workloads.

Qualifications

  • BS or MS in Computer Science or related technical field.
  • Solid understanding of AI/ML inference architecture and the performance characteristics of serving systems.
  • Experience with performance and infrastructure engineering in distributed systems.

Responsibilities

  • Design and build automations to evaluate AI inference performance across hardware generations and configurations.
  • Develop tooling to surface performance trends, regressions, and insights to infrastructure and planning teams.
  • Build projection and forecasting models to support long-term capacity planning decisions.
  • Analyze performance and utilization data to identify bottlenecks, trends, and optimization opportunities.
  • Partner with AI infrastructure engineers, hardware teams, and capacity planners to deliver critical data and tooling.
  • Create and enhance performance analysis workflows to increase team velocity and data reliability.
  • Continuously improve the accuracy, coverage, and usability of performance measurement and analysis systems.

Skills

Python
Go
C++

Education

BS or MS in Computer Science

Tools

Nsight
Prometheus
Grafana

Job description

Apple seeks an engineer to build tooling, automation, and analysis capabilities for the AI inference platform, focusing on benchmarking, capacity projection models, and data pipelines that inform infrastructure teams and capacity planners.

You will design real-time performance dashboards, collaborate with AI infrastructure engineers, hardware teams, and capacity planners to forecast needs months in advance, and help accelerate data-driven decisions for scaling AI workloads.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Inference Platform Performance Engineer
AI Inference Platform Performance Engineer

Socket.dev • Seattle (WA)

On-site
USD 180,000 - 240,000
AI Inference Platform Engineer
AI Inference Platform Engineer

Socket.dev • Seattle (WA)

On-site
USD 180,000 - 240,000
Sr. AI Inference Platform Engineer
Sr. AI Inference Platform Engineer

Apple • Seattle (WA)

On-site
USD 175,000 - 309,000
Employee stock programs
Employee Stock Purchase Plan
Medical and dental coverage
+4
Senior Backend Platform Engineer — AI Tooling & Scale
Senior Backend Platform Engineer — AI Tooling & Scale

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 185,000 - 325,000
Senior AI/ML System Performance Engineer – Architecture
Senior AI/ML System Performance Engineer – Architecture

Socket.dev • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Applied AI Inference Engineer - Benchmark & Optimize
Applied AI Inference Engineer - Benchmark & Optimize

CoreWeave • Bellevue (WA)

On-site
USD 188,000 - 275,000
Medical Insurance
Dental Insurance
Vision Insurance
+12
Inference Performance Engineer - Benchmark & Optimize
Inference Performance Engineer - Benchmark & Optimize

CoreWeave • Sunnyvale (CA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
401(k) with employer match
Paid parental leave
+2
Staff, AI Inference Benchmarking (Equity, Onsite SF)
Staff, AI Inference Benchmarking (Equity, Onsite SF)

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 120,000 - 180,000
Equity
AI Benchmark Architect—Systems & GPUs
AI Benchmark Architect—Systems & GPUs

Sandisk • California (MO)

On-site
USD 120,000 - 160,000
Comprehensive benefits package
Paid vacation and sick leave
Employee Stock Purchase Plan
Inference Performance Engineer: Benchmark & Optimize
Inference Performance Engineer: Benchmark & Optimize

Coreweave • Bellevue (CA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
Equity awards
401(k) with match
+1