Distillation Architect — AI Model Efficiency Lead

Waabi

United States

On-site

USD 195,000 - 286,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive compensation and equity
Health and Wellness benefits (Medical,
Unlimited Vacation
Flexible hours and Work from Home
Daily drinks, snacks and catered meals
Team building activities

Job summary

Waabi is seeking a Distillation Lead to own strategy and execution for distillation across Waabi's AI stack. You will partner with ML Platform, Infrastructure, Onboard Autonomy, and Simulation teams to deliver compressed models meeting latency and memory targets in various deployment contexts.

You will design state‑of‑the‑art distillation pipelines, including diffusion models, QAT/PTQ, and knowledge distillation, while mentoring researchers and engineers and contributing to publications and

Qualifications

  • Strong hands-on distillation, quantization, pruning, and model compression for large neural nets.
  • Bachelor's or Master's in ML, CV, robotics or related field or equivalent experience.
  • Expert Python and PyTorch (or JAX) with experience in distributed training.

Responsibilities

  • Define the technical strategy for model distillation and compression across Waabi's AI stack.
  • Design and scale distillation and efficiency pipelines (diffusion, QAT/PTQ, knowledge distillation).
  • Collaborate with ML Platform, Infrastructure, Onboard Autonomy, and Simulation teams to deploy compressed models.
  • Define benchmarks to evaluate efficiency vs. quality across models and hardware targets.
  • Mentor researchers and engineers in distillation and efficiency; lead technical reviews and talks.

Skills

Distillation
Quantization
Pruning
Model compression
Knowledge distillation
Python & PyTorch

Education

Bachelor's or Master's in ML/CS/Robotics

Tools

Python
PyTorch
JAX
Distributed training

Job description

Waabi is seeking a Distillation Lead to own strategy and execution for distillation across Waabi's AI stack. You will partner with ML Platform, Infrastructure, Onboard Autonomy, and Simulation teams to deliver compressed models meeting latency and memory targets in various deployment contexts.

You will design state‑of‑the‑art distillation pipelines, including diffusion models, QAT/PTQ, and knowledge distillation, while mentoring researchers and engineers and contributing to publications and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Model Distillation & Efficiency Lead
AI Model Distillation & Efficiency Lead

Waabi • San Francisco (CA), Northern (KY)

Hybrid
USD 195,000 - 286,000
Competitive compensation and equity
Health and Wellness benefits (Medical,
Unlimited Vacation
+4
Distillation Lead
Distillation Lead

Waabi • San Francisco (CA), Northern (KY)

Hybrid
USD 195,000 - 286,000
Competitive compensation and equity
Health and Wellness benefits (Medical,
Unlimited Vacation
+4
Ai Distilation Expert
Ai Distilation Expert

RB Labs • San Francisco (CA)

On-site
USD 130,000 - 180,000
AI-Native Backend Engineer
AI-Native Backend Engineer

Distyl • New York (NY), San Francisco (CA)

Hybrid
USD 150,000 - 250,000
Equity
Medical insurance
Dental & Vision
+3
GenAI Research Engineer: Distillation & Efficient Models
GenAI Research Engineer: Distillation & Efficient Models

ByteDance • Seattle (WA)

On-site
USD 242,000 - 456,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+6
Efficient GenAI Research Engineer: Distillation
Efficient GenAI Research Engineer: Distillation

ByteDance • San Jose (CA)

On-site
USD 254,000 - 588,000
Lead ML Training Optimizer for Scalable AI Systems
Lead ML Training Optimizer for Scalable AI Systems

3M HEALTHCARE • San Francisco (CA), Phoenix (AZ), Pittsburgh, Dallas (TX)

Hybrid
USD 140,000 - 210,000
Competitive compensation
Enterprise AI Benchmarking Architect & Evaluation Lead
Enterprise AI Benchmarking Architect & Evaluation Lead

Distyl • San Francisco (CA)

Hybrid
USD 150,000 - 250,000
Equity
Medical, dental, vision
Flexible time off
+4
Applied AI Manager: Distilled Models
Applied AI Manager: Distilled Models

United States Digital Space LLC • Paris (TX)

On-site
USD 150,000 - 210,000
Hybrid work model
AI Data Pipelines Engineer - Labelling & Automation
AI Data Pipelines Engineer - Labelling & Automation

ProducePay • United States

Hybrid
USD 127,000 - 225,000
Competitive compensation
Equity awards
Health benefits
+4