ML Model Serving & High-Performance API Engineer

Black Forest Labs Inc.

San Francisco (CA)

On-site

USD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Travel reimbursement

Job summary

Black Forest Labs Inc. is seeking a Member of Technical Staff to bridge research breakthroughs and production systems.

You will own the bridge between research checkpoints and production-ready inference services, designing scalable APIs and high-performance serving. You will optimize GPU workloads, manage distributed systems, and collaborate with researchers to ship demos and live endpoints rapidly, across cloud infrastructures.

Qualifications

  • Experience building and operating ML inference services in production.
  • Design scalable API architectures with async processing.
  • Optimizing GPU workloads (batching, quantization, compilation, CUDA).
  • Managing distributed systems and task queues under variable load.
  • Implementing monitoring and observability for production ML systems.
  • Debugging performance bottlenecks across model, infrastructure, and network layers.

Responsibilities

  • Turn research checkpoints into production-ready inference services.
  • Design and maintain high-performance APIs serving millions of requests.
  • Optimize inference latency and throughput across GPU infrastructure.
  • Build scalable serving architectures that handle unpredictable traffic.
  • Improve reliability, monitoring, and observability across model-serving systems.
  • Prototype and ship demos that showcase new capabilities in days, not weeks.

Skills

API design
Distributed systems
Performance optimization
Async processing
Cloud platforms

Tools

Docker
Kubernetes
CUDA
Redis
Postgres
AWS
GCP
Azure

Job description

Black Forest Labs Inc. is seeking a Member of Technical Staff to bridge research breakthroughs and production systems.

You will own the bridge between research checkpoints and production-ready inference services, designing scalable APIs and high-performance serving. You will optimize GPU workloads, manage distributed systems, and collaborate with researchers to ship demos and live endpoints rapidly, across cloud infrastructures.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Model Serving & API Backend Engineer
Senior Model Serving & API Backend Engineer

Black Forest Labs • San Francisco (CA)

Hybrid
USD 180,000 - 300,000
Staff Engineer: Large-Scale ML Training Systems
Staff Engineer: Large-Scale ML Training Systems

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 180,000 - 290,000
Equity
Travel cost coverage
Hybrid/onsite collaboration
Member of Technical Staff - Model Serving / API Backend Engineer
Member of Technical Staff - Model Serving / API Backend Engineer

Black Forest Labs • San Francisco (CA)

Hybrid
USD 180,000 - 300,000
Research Engineer: Large-Scale AI Training Systems
Research Engineer: Large-Scale AI Training Systems

BlackForestLabs • San Francisco (CA)

Hybrid
USD 180,000 - 290,000
Equity
Remote Model Serving Engineer for High-Scale ML
Remote Model Serving Engineer for High-Scale ML

Bright Vision Technologies • Canton Charter Township (MI)

On-site
USD 74,000 - 98,000
Model API Engineer - High-Performance Inference
Model API Engineer - High-Performance Inference

Jobzhr • New York (NY)

On-site
USD 180,000 - 360,000
Competitive compensation with equity
100% medical, dental, vision for you +
Flexible PTO including Winter Break
+4
Inference Performance Engineer: Optimize Model Serving
Inference Performance Engineer: Optimize Model Serving

Adaption • San Francisco (CA)

On-site
USD 180,000 - 240,000
Lunch stipend
Travel stipend (Adaption Passport)
Well-being benefits
+1
Member of Technical Staff, Inference & Serving
Member of Technical Staff, Inference & Serving

Inception • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 170,000 - 240,000
Staff Engineer – Foundation Model Serving & APIs
Staff Engineer – Foundation Model Serving & APIs

RoShay Services • San Francisco (CA)

On-site
USD 180,000 - 240,000