Distillation Lead

Waabi

San Francisco, Northern (CA, KY)

Hybrid

USD 195,000 - 286,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive compensation and equity
Health and Wellness benefits (Medical,
Unlimited Vacation
Flexible hours and Work from Home
Daily drinks, snacks and meals
Team building events
Open-ended growth

Job summary

Waabi is seeking a Distillation Lead to own strategy and execution for model distillation across its AI stack, ensuring compressed models meet real-time onboard and high-throughput simulation needs. You will lead partnerships with ML Platform, Infrastructure, Onboard Autonomy, and Simulation teams to deliver efficient models.

You will mentor researchers and engineers, champion best practices for model compression, and stay at the cutting edge of efficiency research with publications and

Qualifications

  • Hands-on experience designing and implementing distillation, quantization, pruning, and model compression techniques for large-scale neural networks.
  • Expert Python and PyTorch (or JAX) skills with experience in large-scale distributed training.
  • Technical leadership: setting direction and driving projects from concept to production.
  • Cross-functional collaboration with infrastructure, platform and autonomy teams.
  • Clear communication of complex trade-offs to diverse audiences.

Responsibilities

  • Define and drive the technical strategy for model distillation and compression across Waabi's AI stack — spanning perception, world models, and planning.
  • Design, implement, and scale distillation and efficiency pipelines for onboard deployment and simulation use-cases.
  • Collaborate with ML Platform, Infrastructure, Onboard, Autonomy, and Simulation teams to integrate compressed models into production pipelines.
  • Define benchmarks and evaluation frameworks to characterize efficiency vs. quality across models and hardware targets.
  • Mentor researchers and engineers, setting a high technical bar and fostering rigorous experimentation.
  • Disseminate knowledge through design reviews, documentation, and internal talks.

Skills

Distillation
Quantization
Pruning
Model compression
Python
PyTorch
Leadership
Cross-functional collaboration

Education

Master's or PhD in ML/CV/Robotics

Tools

TensorRT
ONNX
CUDA

Job description

Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class team, we're unlocking the next era of autonomous transportation with technology that's powering commercial autonomous trucks and robotaxis. Waabi is backed by and partners with world leaders in AI, automotive, logistics, and deep tech.

With offices in Toronto, San Francisco, Dallas, and Pittsburgh, Waabi is growing quickly and looking for diverse, innovative and collaborative candidates who want to impact the world in a positive way. To learn more visit: www.waabi.ai

Waabi's Physical AI platform is powered by state of the art ML models which must be deployed efficiently across diverse use-cases, from onboard vehicle inference to large-scale simulation. As the Distillation Lead, you will own the strategy and execution for distillation across Waabi's AI stack, ensuring our most capable models run efficiently in every deployment context. You will partner closely with ML Platform, Infrastructure, Onboard Autonomy, and Simulation teams to deliver compressed models that meet the performance requirements of both real-time onboard systems and high-throughput simulation pipelines.

You will
  • Define and drive the technical strategy for model distillation and compression across Waabi's AI stack — spanning perception, world models, and planning — with an eye toward both onboard deployment and simulation use-cases.
  • Design, implement, and scale state-of-the-art distillation and efficiency pipelines, which may include:
    • Distillation for generative models (diffusion, autoregressive, flow-matching, video models)
    • Quantization-aware training (QAT) and post-training quantization (PTQ)
    • Knowledge distillation (feature-level, response-based, and relation-based)
    • Structured and unstructured pruning and sparsification
    • Low-rank factorization and efficient architecture design
    • Speculative decoding and other inference-time efficiency techniques
  • Collaborate closely with ML Platform, Infrastructure, Onboard, Autonomy, and Simulation teams to integrate compressed models into production pipelines and meet latency, memory, and throughput targets across deployment contexts.
  • Define rigorous benchmarks and evaluation frameworks to characterize efficiency vs. quality trade-offs across models and hardware targets.
  • Mentor and guide researchers and engineers working in the distillation and model efficiency space, setting a high technical bar and fostering a culture of rigorous experimentation.
  • Champion best practices for model compression across the organization; disseminate knowledge through internal design reviews, documentation, and technical talks.
  • Stay at the cutting edge of model efficiency research; contribute to the broader scientific community through publications and open-source contributions.
Qualifications
  • Deep distillation expertise: You have extensive hands-on experience designing and implementing distillation, quantization, pruning, and model compression techniques for large-scale neural networks, with demonstrated impact in production settings.
  • Strong research and engineering foundation: A Bachelor's or Master's degree in Machine Learning, Computer Vision, Robotics, or a related field, or equivalent industry experience; relevant hands-on experience in model distillation and efficiency is what matters most. Expert Python and PyTorch (or JAX) skills with experience in large-scale distributed training.
  • Technical leadership: You have a proven track record of setting technical direction and driving projects from conception to production. You inspire and elevate those around you through deep technical expertise and mentorship.
  • Cross-functional collaboration: You have experience working closely with infrastructure, platform, and autonomy teams to deploy compressed models under real engineering constraints.
  • Clear communicator: You can communicate complex technical trade-offs clearly to diverse audiences and drive alignment across research and engineering teams.
Bonus
  • Experience with hardware-aware optimization (TensorRT, ONNX, custom CUDA kernels, hardware-specific quantization).
  • Publications at top-tier ML/CV venues (NeurIPS, ICML, CVPR, ICLR, ECCV) in model compression, efficient deep learning, or related areas.
  • Experience distilling large generative models (diffusion models, LLMs, VLMs, or video models).
  • Background in autonomous vehicles or robotics.

The US yearly salary range for this role is: $195,000 - $286,000 USD in addition to competitive perks & benefits. Waabi (US) Inc.'s yearly salary ranges are determined based on several factors in accordance with the Company's compensation practices. The salary base range is reflective of the minimum and maximum target for new hire salaries for the position across all US locations. Note: The Company provides additional compensation for employees in this role, including equity incentive awards and an annual performance bonus.

Perks/Benefits
  • Competitive compensation and equity awards.
  • Health and Wellness benefits encompassing Medical, Dental and Vision coverage (for full-time employees only).
  • Unlimited Vacation.
  • Flexible hours and Work from Home support.
  • Daily drinks, snacks and catered meals (when in office).
  • Regularly scheduled team building activities and social events both on-site, off-site & virtually.
  • As we grow, this list continues to evolve!

Waabi is a technology start-up building technologies to transform the way the world moves. Join our talented team to be a part of the future and to make an impact!

Waabi is an equal opportunity employer. We celebrate diversity and are committed to creating a supportive, inclusive, and accessible workplace for all our employees. We seek applicants of all backgrounds and identities, across race, color, ethnicity, national origin or ancestry, age, citizenship, religion, sex, sexual orientation, gender identity or expression, military or veteran status, marital status, pregnancy or parental status, caregiver status, disability, or any other characteristic protected by law. We make workplace accommodations for qualified individuals with disabilities as required by applicable law. If reasonable accommodation is needed to participate in the job application or interview process please let our recruiting team know.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Distillation Lead
Distillation Lead

Waabi • United States

On-site
USD 195,000 - 286,000
Competitive compensation and equity
Health and Wellness benefits (Medical,
Unlimited Vacation
+3
Senior / Staff Software Engineer, ML Datasets & Data Pipelines
Senior / Staff Software Engineer, ML Datasets & Data Pipelines

ProducePay • United States

On-site
USD 148,000 - 260,000
Competitive compensation and equity
Health and Wellness benefits (Medical,
Unlimited Vacation
+3
Software Engineer, Labelling, Data & Automation
Software Engineer, Labelling, Data & Automation

ProducePay • United States

Hybrid
USD 127,000 - 225,000
Competitive compensation
Equity awards
Health benefits
+4
Software Engineer, Labelling, Data & Automation
Software Engineer, Labelling, Data & Automation

Waabi • San Francisco (CA)

On-site
USD 127,000 - 225,000
Competitive compensation
Health and Wellness benefits
Unlimited Vacation
+3
Senior / Staff Research Engineer, Simulation Assets & Content Systems
Senior / Staff Research Engineer, Simulation Assets & Content Systems

ProducePay • United States

Hybrid
USD 155,000 - 269,000
Health and wellness benefits
Unlimited vacation
Flexible hours and work from home
+2
Senior ML Systems Engineer — Auto Labelling & Perception
Senior ML Systems Engineer — Auto Labelling & Perception

Waabi • Dallas (TX)

On-site
USD 170,000 - 220,000
Competitive compensation and equity awards
Health and Wellness benefits
Unlimited Vacation
+3
Senior / Staff Software Engineer, Simulation Platform
Senior / Staff Software Engineer, Simulation Platform

ProducePay • United States

Hybrid
USD 159,000 - 268,000
Equity awards
Health benefits
Unlimited vacation
+2
Research Engineer, Neural Rendering
Research Engineer, Neural Rendering

ProducePay • United States

Hybrid
USD 134,000 - 235,000
Competitive compensation and equity
Health and Wellness benefits
Unlimited Vacation
+4
Research Engineer, World Models
Research Engineer, World Models

Waabi • San Francisco (CA)

On-site
USD 155,000 - 269,000
Competitive compensation and equity awards
Health and Wellness benefits
Unlimited Vacation
+3
Senior / Staff Software Engineer, Mapping
Senior / Staff Software Engineer, Mapping

ProducePay • United States

On-site
USD 141,000 - 242,000
Equity awards
Health benefits
Unlimited vacation
+3