ML Engineer: Production-Scale LLMs & AI Infra

Harrison Clarke

San Francisco (CA)

On-site

USD 150,000 - 190,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Harrison Clarke is seeking a technically talented engineer to join a well-funded AI startup in San Francisco, building cutting-edge ML systems at the intersection of data pipelines, model training, and production inference. You will own the full ML lifecycle—from data generation and training to deployment and monitoring, collaborating with researchers to turn ideas into reliable production systems.

A strong foundation in Python and ML frameworks, plus knowledge of transformers, is required, with

Qualifications

  • Bachelor's or Master's in CS or a related technical discipline.
  • Strong Python programming skills and experience with modern ML frameworks.
  • Solid understanding of transformer architectures and large language models.
  • Experience building production-quality ML systems.
  • Comfortable working across both research and engineering environments.
  • Strong software engineering fundamentals and systems thinking.

Responsibilities

  • Building scalable data pipelines to collect, process, and generate large synthetic datasets for ML.
  • Profiling and optimizing model training and inference performance.
  • Deploying and maintaining high-throughput inference systems for LLMs.
  • Working closely with researchers to translate new ideas into reliable production systems.
  • Building tooling that supports the ML development lifecycle from experimentation through deployment.
  • Contributing to technical research, experimentation, and engineering best practices.

Skills

Python programming
ML frameworks
Transformer architectures
Production ML systems
Software engineering fundamentals
Research & engineering collaboration

Education

Bachelors/Masters in CS or related

Tools

DeepSpeed
FSDP
Ray
Spark
Beam

Job description

Harrison Clarke is seeking a technically talented engineer to join a well-funded AI startup in San Francisco, building cutting-edge ML systems at the intersection of data pipelines, model training, and production inference. You will own the full ML lifecycle—from data generation and training to deployment and monitoring, collaborating with researchers to turn ideas into reliable production systems.

A strong foundation in Python and ML frameworks, plus knowledge of transformers, is required, with

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML-Driven Infra Engineer: Scale Data Pipelines & Models
ML-Driven Infra Engineer: Scale Data Pipelines & Models

Harrison Clarke • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior LLM Infra Engineer Scalable AI Agents
Senior LLM Infra Engineer Scalable AI Agents

Harrison Clarke • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation package
Significant technical ownership
Meaningful equity
ML Software Engineer — AI Safety & Production NLP
ML Software Engineer — AI Safety & Production NLP

Harrison Clarke • San Francisco (CA)

On-site
USD 140,000 - 210,000
Founding ML Infra Engineer — Production-Grade LLMs
Founding ML Infra Engineer — Production-Grade LLMs

Realmlabs • Sunnyvale (CA)

On-site
USD 210,000 - 350,000
Market aligned compensation
Founding engineer equity
Medical, Dental, Vision, and Life insurance
+2
Production ML Engineer — Scale & Train LLMs
Production ML Engineer — Scale & Train LLMs

Anthropic • San Francisco (CA)

On-site
USD 350,000 - 850,000
Generous vacation and parental leave
Flexible working hours
Office space for collaboration
Applied AI Engineer - LLM Orchestration & Production AI
Applied AI Engineer - LLM Orchestration & Production AI

Radley James • San Francisco (CA)

On-site
USD 180,000 - 275,000
Senior AI Engineer: Build Production-Grade LLM Solutions
Senior AI Engineer: Build Production-Grade LLM Solutions

Innovaccer Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Generous PTO: 20 days per year
Parental Leave
Rewards & Recognition program
+1
ML Engineer — NLP & Agentic AI for Production
ML Engineer — NLP & Agentic AI for Production

Ema • San Francisco (CA)

On-site
USD 160,000 - 220,000
AI Engineer - Build Production-Grade LLM Systems
AI Engineer - Build Production-Grade LLM Systems

Innovaccer Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Generous PTO
Parental Leave
Rewards & Recognition
+1
Senior AI Systems Architect - Production LLM & RAG
Senior AI Systems Architect - Production LLM & RAG

Harnham • San Francisco (CA)

On-site
USD 150,000 - 200,000