Founding Machine Learning Engineer

Velvet

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A data research company is seeking a Founding Machine Learning Engineer in San Francisco to build and enhance ML pipelines for processing audio and video data. This role combines hands-on execution with engineering challenges. Responsibilities include designing processing scripts and fine-tuning open-source models. Ideal candidates should have strong experience in ML infrastructure and proficiency in PyTorch. The position requires adaptability and a commitment to high data quality, impacting the performance of AI models.

Qualifications

  • Strong experience in ML infrastructure, speech/audio processing, or large-scale data pipelines.
  • Proficiency in PyTorch with familiarity in distributed job orchestration.
  • Ability to work effectively in an early-stage environment where scope is broad.

Responsibilities

  • Build and enhance post-processing pipelines for large volumes of video and audio data.
  • Deploy and fine-tune open-source models for speech recognition and related tasks.
  • Design infrastructure for large-scale distributed processing across cloud platforms.

Skills

ML infrastructure
Speech/audio processing
Large-scale data pipelines
Proficiency in PyTorch

Job description

About Us

Velvet is a data research company building the datasets that power the next generation of multimodal AI. Founded by Lucas Mantovani (ex Meta FAIR) and Lucas Tucker (ex Adobe Infrastructure), our mission is to make AI more human by producing high-quality audiovisual training data for frontier labs.

We're hiring a Founding Machine Learning Engineer to build the pipelines that turn raw footage into clean, structured training data. This is a hands‑on, execution‑heavy role at the intersection of ML engineering and research. You'll own the full lifecycle — from writing and testing processing scripts to deploying them at scale across thousands of hours of video.

What You'll Do
  • Build and enhance post‑processing pipelines that clean, validate, and package large volumes of video and audio data for multimodal model training. These pipelines must handle wide variation in speech, visual quality, and format — making robustness a huge engineering challenge.
  • Deploy and fine‑tune open‑source models for speech recognition, speaker diarization, video segmentation, and related tasks.
  • Design infrastructure for large‑scale distributed processing — parallelizing thousands of compute jobs across cloud platforms and optimizing for throughput and cost.
What We're Looking For
  • Strong experience in ML infrastructure, speech/audio processing, or large-scale data pipelines.
  • Proficiency in PyTorch. Familiarity with distributed job orchestration.
  • Claude Code pilled.
  • A bias toward shipping. You default to building, not theorizing.
  • Ability to work effectively in an early‑stage environment where scope is broad and priorities shift fast.
Even Better
  • Prior work at a data company or frontier AI lab.
  • Track record building pipelines that process tens of thousands of hours of audio or video.
  • Experience with infrastructure cost optimization or model fine‑tuning for production use.
You'll Thrive Here If
  • You're energized by operational work with immediate, visible impact.
  • You treat broken processes as engineering problems worth solving properly.
  • You hold yourself to a high bar for data quality — because you understand it directly determines model performance.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist
Research Scientist

Velvet • San Francisco (CA)

On-site
USD 110,000 - 150,000
Member of Technical Staff, Machine Learning - NomadicML
Member of Technical Staff, Machine Learning - NomadicML

Praxis, Inc. • San Francisco (CA)

On-site
USD 150,000 - 240,000
Member of Technical Staff, Machine Learning
Member of Technical Staff, Machine Learning

Nomadic AI • San Francisco (CA)

On-site
USD 170,000 - 250,000
Member of Technical Staff, Machine Learning - NomadicML
Member of Technical Staff, Machine Learning - NomadicML

Pear VC • Austin (TX), California (MO)

On-site
USD 120,000 - 150,000
Founding ML Engineer
Founding ML Engineer

Crustdata (YC F24) • San Francisco (CA)

On-site
USD 150,000 - 200,000
Machine Learning Engineer, Applied AI Infrastructure
Machine Learning Engineer, Applied AI Infrastructure

Bonfirevc • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Member of Technical Staff - Machine Learning Infrastructure Engineer, Post-training
Member of Technical Staff - Machine Learning Infrastructure Engineer, Post-training

Preference Model • Seattle (WA)

On-site
USD 180,000 - 300,000
Health insurance
Vision insurance
Dental insurance
+3
Machine Learning Engineer
Machine Learning Engineer

ExaCare AI • New York (NY)

Hybrid
USD 110,000 - 150,000
Member of Technical Staff - Machine Learning Infrastructure Engineer, Post-training
Member of Technical Staff - Machine Learning Infrastructure Engineer, Post-training

Preference Model • San Francisco (CA)

On-site
USD 180,000 - 300,000
Cash and equity
Ownership & autonomy
Visa sponsorship
+5
ML Engineer
ML Engineer

Catalyst Labs • Georgia

On-site
USD 80,000 - 100,000