Senior ML Infrastructure Engineer — GPU & Robotics

Nimble

San Francisco (CA)

On-site

USD 210,000 - 300,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Unlimited Flexible Time Off
Health Insurance
Paid Parental Leave
Commuter Benefits
Referral Bonus
401k
Equity

Job summary

Nimble is seeking a Software Engineer to join the ML Infrastructure team. You will help build the training and inference systems that power our general-purpose warehouse robots.

You’ll own training infrastructure end to end, keeping GPUs highly utilized and enabling researchers to launch new experiments with a single command. You’ll collaborate with ML and Robotics teams to design scalable, high-throughput platforms for model development.

Qualifications

  • BA/BS/MS/PhD in CS or related field, or equivalent practical experience.
  • 4+ years in infrastructure, distributed systems, ML systems, robotics, or related area.
  • Experience with Rust, Go, Python, or C++.
  • Experience with PyTorch or JAX.
  • Strong understanding of distributed systems and performance optimization.
  • Experience with Kubernetes orchestration and large distributed jobs.
  • Ability to optimize GPU memory hierarchy and multi-GPU operations.

Responsibilities

  • Design, develop, and maintain ML training infrastructure enabling AI teams to run training jobs efficiently.
  • Build low-latency inference pipelines for production robotics workloads.
  • Develop, tune, and optimize CUDA kernels.
  • Design scalable training-platform systems for high-throughput data ingestion and distributed training.
  • Participate in design reviews and evaluate technical tradeoffs.

Skills

Rust
Go
Python
C++
PyTorch
JAX
Distributed systems
Kubernetes
GPU memory tuning
Systems optimization
Cross-functional collaboration

Education

Bachelor's, Master's, or PhD in Computer Science or related field

Tools

Kubernetes
PyTorch
JAX
DeepSpeed
Megatron
NCCL
Parquet
Arrow

Job description

Nimble is seeking a Software Engineer to join the ML Infrastructure team. You will help build the training and inference systems that power our general-purpose warehouse robots.

You’ll own training infrastructure end to end, keeping GPUs highly utilized and enabling researchers to launch new experiments with a single command. You’ll collaborate with ML and Robotics teams to design scalable, high-throughput platforms for model development.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer: ML Infra
Software Engineer: ML Infra

Generalist • Somerville (MA), San Mateo (CA)

On-site
USD 120,000 - 160,000
Senior ML Training Infrastructure Engineer
Senior ML Training Infrastructure Engineer

Dyna Robotics, Inc • Redwood City (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
ML Infra Engineer: GPU Fleet & Inference Orchestrator
ML Infra Engineer: GPU Fleet & Inference Orchestrator

Generalist • San Francisco (CA)

On-site
USD 120,000 - 160,000
Software Engineer, ML Infrastructure
Software Engineer, ML Infrastructure

Cursor • New York (NY), San Francisco (CA)

On-site
USD 120,000 - 150,000
Senior Remote ML Infrastructure Engineer: GPU & Scale
Senior Remote ML Infrastructure Engineer: GPU & Scale

Bright Vision Technologies • Bellevue (WA)

On-site
USD 100,000 - 150,000
Systems Engineer - Machine Learning
Systems Engineer - Machine Learning

General Robotics • Redmond (WA)

On-site
USD 155,000 - 200,000
Medical benefits
401K
Senior Software Engineer - ML Infrastructure
Senior Software Engineer - ML Infrastructure

Claryo • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Medical/Dental/Vision
401k with employer matching
Parental leave
+1
Senior ML Infra Engineer - Scale GPU Clusters, Remote
Senior ML Infra Engineer - Scale GPU Clusters, Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 500,000
Equity
Medical/Dental/Vision coverage
Unlimited PTO
+1
Software Engineer: ML Infra
Software Engineer: ML Infra

Generalist • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior ML Infra Engineer — Scale GPU Clusters + Equity
Senior ML Infra Engineer — Scale GPU Clusters + Equity

Nuro • Mountain View (CA)

On-site
USD 193,000 - 292,000
Annual performance bonus
Equity options
Comprehensive benefits package