Member of Technical Staff - Model Serving / API Backend Engineer

Headline - Asia

San Francisco (CA)

Hybrid

USD 180,000 - 300,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Black Forest Labs seeks a Backend/ML Platform engineer to bridge research breakthroughs and production systems. You will create production-ready inference services with fast, scalable APIs and optimize GPU infrastructure for low latency.

You’ll collaborate with researchers to move ideas to live endpoints and build robust, observable model-serving stacks. You will work in a distributed setup with offices in Freiburg or SF or remote arrangements, with periodic in-person weeks and travel coverage

Qualifications

  • Built and operated systems at meaningful scale.
  • Experience scaling APIs or ML systems under load.
  • Comfort working in fast-moving, research-adjacent environments.

Responsibilities

  • Turn research checkpoints into production-ready inference services.
  • Design and maintain high-performance APIs serving millions of requests.
  • Optimize inference latency and throughput across GPU infrastructure.
  • Build scalable serving architectures for unpredictable traffic.
  • Improve reliability, monitoring, and observability across model-serving systems.
  • Prototype and ship demos showing new capabilities quickly.
  • Collaborate with researchers to move ideas to live endpoints.

Skills

API design
Distributed systems
Performance optimization
Cost-aware engineering
Research-to-production bridge

Tools

Python
FastAPI
async systems
GPU infrastructure
CUDA
Docker
Kubernetes
Redis
Postgres
AWS
GCP
Azure
Observability stacks

Job description

About Black Forest Labs

We're the team behind Latent Diffusion, Stable Diffusion, and FLUX—foundational technologies that changed how the world creates images and video. We’re creating the generative models that power how people make images and video—tools used by millions of creators, developers, and businesses worldwide. Our FLUX models are among the most advanced in the world, and we're just getting started.

About Black Forest Labs

We're the team behind Latent Diffusion, Stable Diffusion, and FLUX—foundational technologies that changed how the world creates images and video. We’re creating the generative models that power how people make images and video—tools used by millions of creators, developers, and businesses worldwide. Our FLUX models are among the most advanced in the world, and we're just getting started.

Headquartered in Freiburg, Germany with a growing presence in San Francisco, we're scaling fast while staying true to what makes us different: research excellence, open science, and building technology that expands human creativity.

Why This Role

Our research team moves fast. Models improve weekly. New capabilities emerge constantly.

What slows us down is not model quality—it’s productionization.

Without This Role
  • Research checkpoints sit longer before becoming usable APIs
  • Inference is slower than it needs to be
  • APIs struggle under load
  • Demos don’t reflect the true potential of our models

This role removes the bottleneck between frontier research and production reality. Once hired, researchers ship faster, demos launch faster, and customers experience models at their best.

What You’ll Work On

You will own the bridge between research breakthroughs and production systems.

  • Turn research checkpoints into production-ready inference services
  • Design and maintain high-performance APIs serving millions of requests
  • Optimize inference latency and throughput across GPU infrastructure
  • Build scalable serving architectures that handle unpredictable traffic
  • Improve reliability, monitoring, and observability across model-serving systems
  • Prototype and ship demos that showcase new capabilities in days, not weeks
  • Collaborate closely with researchers to move from idea to live endpoint rapidly
Tools & Context – Model Serving & API Infrastructure
  • Python, FastAPI, async systems
  • GPU infrastructure, CUDA, inference optimization
  • Docker and Kubernetes
  • Redis, Postgres, distributed task queues
  • Cloud platforms (AWS, GCP, or Azure)
  • Observability stacks (metrics, logging, tracing)

This role spans backend systems, GPU performance, and production ML serving.

What We’re Looking For

You’ve built and operated systems at meaningful scale. You understand the difference between a research prototype and a production system. You are comfortable navigating ambiguity, making tradeoffs, and improving systems under real-world constraints.

You Demonstrate
  • Strong judgment around performance, reliability, and cost tradeoffs
  • Experience scaling APIs or ML systems under load
  • Comfort working in fast-moving, research-adjacent environments
  • Ownership from system design through debugging and deployment
Role-specific Experience We Value
  • Building and operating ML inference services in production
  • Designing scalable API architectures with async processing
  • Optimizing GPU workloads (batching, quantization, compilation, CUDA)
  • Managing distributed systems and task queues under variable load
  • Implementing monitoring and observability for production ML systems
  • Debugging performance bottlenecks across model, infrastructure, and network layers
Bonus Experience Includes
  • Real-time or low-latency inference systems
  • TensorRT, reduced precision, layer fusion, or model compilation techniques
  • Frontend demo tooling (Streamlit, Gradio, React)
  • CI/CD and automated testing for ML systems
  • Security best practices for API and model serving
How We Work Together

We’re a distributed team with real offices that people actually use. Depending on your role, you’ll either join us in Freiburg or SF at least 2 days a week (or one full week every other week), or work remotely with a monthly in‑person week to stay connected. We’ll cover reasonable travel costs to make this possible. We think in‑person time matters, and we’ve structured things to make it accessible to all. We’ll discuss what this will look like for the role during our interview process.

Everything We Do Is Grounded In Four Values
  • Obsessed. We are a frontier research lab. The science has to be right, the understanding deep, the product beautiful.
  • Low Ego. The work speaks. The best idea wins, no matter who said it. Credit is shared. Nobody is above any task.
  • Bold. We take the ambitious bet. We ship, we do not wait for conditions to be perfect.
  • Kind. People over politics. We treat each other with genuine warmth. Agency without empathy creates chaos.
Base Annual Salary:

$180,000–$300,000 USD

We're based in Europe and value depth over noise, collaboration over hero culture, and honest technical conversations over hype. Our models have been downloaded hundreds of millions of times, but we're still a ~50-person team learning what's possible at the edge of generative AI.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff - Model Serving / API Backend Engineer
Member of Technical Staff - Model Serving / API Backend Engineer

Black Forest Labs • San Francisco (CA)

Hybrid
USD 180,000 - 300,000
Product Engineer
Product Engineer

Black Forest Labs • United States

Hybrid
USD 163,000 - 221,000
Competitive salary
Equity
Benefits
+1
Member of Technical Staff - Research Engineer
Member of Technical Staff - Research Engineer

Black Forest Labs • San Francisco (CA)

On-site
USD 180,000 - 290,000
Senior Partnerships Manager
Senior Partnerships Manager

Black Forest Labs Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 175,000 - 240,000
Senior Partnerships Manager
Senior Partnerships Manager

Black Forest Labs • San Francisco (CA)

Hybrid
USD 135,000 - 240,000
Equity
Remote-friendly
Member of Technical Staff - Research Engineer
Member of Technical Staff - Research Engineer

Black Forest Labs Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 290,000
Senior Partnerships Manager
Senior Partnerships Manager

Black Forest Labs • Seattle (WA)

Hybrid
USD 175,000 - 240,000
Senior Partnerships Manager
Senior Partnerships Manager

Headline - Asia • San Francisco (CA)

Hybrid
USD 175,000 - 240,000
Travel costs covered
Remote-friendly with in-person weeks
IT Engineer
IT Engineer

Black Forest Labs • United States

Hybrid
USD 120,000 - 150,000
Member of Technical Staff, Product Engineering
Member of Technical Staff, Product Engineering

San Francisco Tensor Company • San Francisco (CA)

On-site
USD 225,000 - 275,000
Relocation assistance
Equity