Engineering Manager, ML Platform & GPU Infra

applied

Sunnyvale (CA)

On-site

USD 204,000 - 343,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity options
Health insurance
401k retirement benefits
Paid time off

Job summary

Applied Intuition is seeking an Engineering Manager for the ML Platform team in Sunnyvale, California. This role involves leading a team focused on building and optimizing a cutting-edge infrastructure for Physical AI. You will manage GPU clusters, drive performance optimizations, and collaborate with various teams to accelerate development.

Your responsibilities include team management, infrastructure design, and ensuring operational excellence. Ideal candidates have substantial experience in engineering management and a passion for high-performance systems.

Qualifications

  • 3+ years of engineering management experience, ideally leading infrastructure or platform teams.
  • Deep experience with distributed systems, GPU computing, or large-scale ML infrastructure.
  • Track record of building and operating systems that run reliably at massive scale.

Responsibilities

  • Grow and manage a team of world-class engineers for ML platforms.
  • Own design and evolution of frameworks for orchestrating training jobs.
  • Drive scaling of GPU cluster infrastructure.

Skills

Engineering management experience
Distributed systems
GPU computing
Large-scale ML infrastructure
Team leadership

Education

Bachelor's or higher in relevant field

Tools

PyTorch Distributed
Kubernetes
Slurm
GPU cluster management

Job description

Applied Intuition is seeking an Engineering Manager for the ML Platform team in Sunnyvale, California. This role involves leading a team focused on building and optimizing a cutting-edge infrastructure for Physical AI. You will manage GPU clusters, drive performance optimizations, and collaborate with various teams to accelerate development.

Your responsibilities include team management, infrastructure design, and ensuring operational excellence. Ideal candidates have substantial experience in engineering management and a passion for high-performance systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, GPU ML Accelerators
Engineering Manager, GPU ML Accelerators

SignalAI • New York (NY)

Hybrid
USD 500,000 - 850,000
optional equity donation matching
generous vacation and parental leave
flexible working hours
+1
Senior ML Infra Engineer - End-to-End Pipelines & GPU Training
Senior ML Infra Engineer - End-to-End Pipelines & GPU Training

Applied Intuition • Sunnyvale (CA)

On-site
USD 153,000 - 222,000
Health insurance
401k retirement benefits
Learning stipends
+1
Engineering Manager, AI Inference Platform & ML Infra
Engineering Manager, AI Inference Platform & ML Infra

Google • Sunnyvale (CA)

On-site
USD 207,000 - 301,000
ML Performance Engineer: Scale GPU-Driven Training
ML Performance Engineer: Scale GPU-Driven Training

Decisive Point • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads

VeeAR Projects Inc. • Sunnyvale (CA)

On-site
USD 80,000 - 100,000
Engineering Manager - ML Platform and Infrastructure
Engineering Manager - ML Platform and Infrastructure

applied • Sunnyvale (CA)

On-site
USD 204,000 - 343,000
Equity options
Health insurance
401k retirement benefits
+1
Engineering Manager, GPU ML Accelerator
Engineering Manager, GPU ML Accelerator

Menlo Ventures • New York (NY)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
Senior Remote ML Infrastructure Engineer: GPU & Scale
Senior Remote ML Infrastructure Engineer: GPU & Scale

Bright Vision Technologies • Bellevue (WA)

On-site
USD 100,000 - 150,000
Engineering Manager, GPU Infrastructure & Platforms
Engineering Manager, GPU Infrastructure & Platforms

cohere • United States

Hybrid
USD 180,000 - 240,000
Lunch stipend
Health & dental benefits
RRSP/401K matching
+5
Engineering Manager - ML Infrastructure & Systems
Engineering Manager - ML Infrastructure & Systems

Thinking Machines Lab • San Francisco (CA)

On-site
USD 400,000 - 500,000
Generous health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1