AI Inference Platform Performance Engineer

Socket.dev

Seattle (WA)

On-site

USD 180,000 - 240,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple is seeking an engineer to build tooling, automation, and analysis capabilities for its AI inference platform. You will develop advanced performance benchmarking systems, capacity projection models, and data pipelines to inform capacity-planning and infrastructure decisions.

You will work at the intersection of AI systems performance, distributed infrastructure, and software engineering to turn raw data into actionable insight for scaling and optimizing the inference platform.

Qualifications

  • BS or MS in Computer Science or related field.
  • Strong understanding of AI/ML inference architecture and performance characteristics of serving systems.
  • Experience with performance and infrastructure engineering in distributed systems.
  • Proficiency in Python, Go, C++, or other languages.
  • Experience with automation, tooling and data pipelines for engineering workflows.
  • Strong knowledge of GPU/accelerator architecture for AI workloads.
  • Practical statistics applicable to performance analysis and forecasting.
  • Excellent communication skills to guide infrastructure teams.

Responsibilities

  • Build tooling and analysis systems to monitor AI inference performance at scale.
  • Develop benchmarking frameworks and capacity projection models for forecasting needs.
  • Create data pipelines to support engineering workflows and capacity planning.
  • Collaborate with AI infrastructure teams to inform hardware investments and scaling decisions.

Skills

AI/ML inference architecture
Distributed systems performance
Python/Go/C++ programming
Automation tooling
GPU/accelerator architecture
Statistical analysis
Clear communication

Education

BS or MS in Computer Science

Tools

Nsight
Prometheus
Grafana
Kubernetes

Job description

Apple is seeking an engineer to build tooling, automation, and analysis capabilities for its AI inference platform. You will develop advanced performance benchmarking systems, capacity projection models, and data pipelines to inform capacity-planning and infrastructure decisions.

You will work at the intersection of AI systems performance, distributed infrastructure, and software engineering to turn raw data into actionable insight for scaling and optimizing the inference platform.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Inference Platform Engineer — Benchmarking
Senior AI Inference Platform Engineer — Benchmarking

Apple • Seattle (WA)

On-site
USD 175,000 - 309,000
Employee stock programs
Employee Stock Purchase Plan
Medical and dental coverage
+4
AI Inference Platform Engineer
AI Inference Platform Engineer

Socket.dev • Seattle (WA)

On-site
USD 180,000 - 240,000
Sr. AI Inference Platform Engineer
Sr. AI Inference Platform Engineer

Apple • Seattle (WA)

On-site
USD 175,000 - 309,000
Employee stock programs
Employee Stock Purchase Plan
Medical and dental coverage
+4
Senior AI/ML System Performance Engineer – Architecture
Senior AI/ML System Performance Engineer – Architecture

Socket.dev • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
AI-Driven Performance Systems Engineer & Agentic Tools
AI-Driven Performance Systems Engineer & Agentic Tools

Apple Inc. • Cupertino (CA)

On-site
USD 129,000 - 225,000
Inference Performance Engineer - Benchmark & Optimize
Inference Performance Engineer - Benchmark & Optimize

CoreWeave • Sunnyvale (CA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
401(k) with employer match
Paid parental leave
+2
Inference Performance Engineer
Inference Performance Engineer

Neura Market • Bellevue (WA), Northern (KY)

Hybrid
USD 188,000 - 275,000
Medical, dental, and vision insurance
Company-paid Life Insurance
401(k) with generous match
+2
Applied AI Inference Performance Engineer
Applied AI Inference Performance Engineer

CoreWeave • San Francisco (CA)

On-site
USD 188,000 - 275,000
Medical benefits
401(k) with match
Flexible PTO
+5
AI Model Performance Engineer - Apple Silicon
AI Model Performance Engineer - Apple Silicon

Apple Inc. • Cupertino (CA)

On-site
USD 184,000 - 325,000
Employee stock programs
Education reimbursement
Medical and dental coverage