ML Infra Engineer: GPU Fleet & Inference Orchestrator

Generalist

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Generalist is seeking a candidate to manage GPU fleets for training large-scale AI models. You will optimize ML data loading, storage, and orchestration of robot inference fleets in compute-constrained environments. Ideal candidates have deep experience with GPUs, Slurm or Kubernetes, and a strong understanding of the ML hardware stack.

Your role will significantly contribute to making general-purpose robots a reality. Join a team from leading AI labs committed to pioneering robotics and AI advancements.

Qualifications

  • Managed large fleets of GPUs for ML training or inference.
  • Experience in orchestrating workflows using Slurm or Kubernetes.
  • Built systems for high-scale data loading.

Responsibilities

  • Own GPU compute fleets.
  • Ensure usability and maximization of GPU resources.
  • Optimize ML data loading and storage.
  • Orchestrate robot inference fleets.

Skills

GPU management
ML workload orchestration (Slurm/Kubernetes)
High-scale data loading
Understanding of ML hardware stack
Experience with NVidia GPUs

Job description

Generalist is seeking a candidate to manage GPU fleets for training large-scale AI models. You will optimize ML data loading, storage, and orchestration of robot inference fleets in compute-constrained environments. Ideal candidates have deep experience with GPUs, Slurm or Kubernetes, and a strong understanding of the ML hardware stack.

Your role will significantly contribute to making general-purpose robots a reality. Join a team from leading AI labs committed to pioneering robotics and AI advancements.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Infra & GPU Fleet Engineer
ML Infra & GPU Fleet Engineer

Generalist • Somerville (MA), San Mateo (CA)

On-site
USD 120,000 - 160,000
Software Engineer: ML Infra
Software Engineer: ML Infra

Generalist • Somerville (MA), San Mateo (CA)

On-site
USD 120,000 - 160,000
Software Engineer: ML Infra
Software Engineer: ML Infra

Generalist • San Francisco (CA)

On-site
USD 120,000 - 160,000
ML Infra Engineer — GPU Clusters & Distributed Systems
ML Infra Engineer — GPU Clusters & Distributed Systems

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 170,000 - 250,000
Industry-leading compensation and/or:?
Unlimited PTO
Top-tier medical, dental, and vision
+1
ML Infra Engineer: Scale GPU Training & Data Pipelines
ML Infra Engineer: Scale GPU Training & Data Pipelines

Humble Robotics • United States

On-site
USD 150,000 - 230,000
Senior ML Infrastructure Engineer — GPU & Robotics
Senior ML Infrastructure Engineer — GPU & Robotics

Nimble • San Francisco (CA)

On-site
USD 210,000 - 300,000
Unlimited Flexible Time Off
Health Insurance
Paid Parental Leave
+4
Autonomous AI Infrastructure Engineer: GPU Fleet Mastery
Autonomous AI Infrastructure Engineer: GPU Fleet Mastery

Together • San Francisco (CA)

On-site
USD 190,000 - 270,000
Health insurance
Startup equity
Competitive benefits
Senior ML Infra Engineer - Scale GPU Clusters, Remote
Senior ML Infra Engineer - Scale GPU Clusters, Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 500,000
Equity
Medical/Dental/Vision coverage
Unlimited PTO
+1
Senior ML Infra Architect - Data Pipelines & GPUs
Senior ML Infra Architect - Data Pipelines & GPUs

Generalintuition • New York (NY)

On-site
USD 180,000 - 250,000
Remote ML Infra Architect: Scalable GPU Training
Remote ML Infra Architect: Scalable GPU Training

Bright Vision Technologies • Plymouth (MN)

Remote
USD 100,000 - 150,000