ML Inference Engineer

United States Digital Space LLC

Greater London

On-site

GBP 80,000 - 120,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

United States Digital Space LLC is seeking a Research Engineer to own the backend that serves world models. You will build and run the inference systems behind our product and API, turning frontier research models into fast, reliable services.

This is a hands‑on role for someone who has owned production APIs in an AI system and can collaborate with researchers to bring models to users. Responsibilities include owning inference services, optimizing GPU utilization, deploying models as production

Qualifications

  • 3+ years of software engineering in an AI, machine learning, or computer vision product environment.
  • Experience designing and owning production APIs.
  • Strong Python, and experience shipping containerized services on GPU inference platforms.
  • Hands‑on GPU inference performance work: profiling, memory, batching, and cold start.
  • Familiarity with 3D or computer vision data is a strong plus.

Responsibilities

  • Own the inference services and APIs that power our product, from request to generated output.
  • Optimize serving performance, reliability, and cost at scale. Maximize GPU utilization and minimize device memory footprint.
  • Apply latest techniques for improving inference performance.
  • Deploy research models as production services, including customer‑specific configurations.
  • Work with research and product colleagues to bring new models to users.

Skills

Production APIs
Python
Containerized services
GPU inference optimization
GPU performance profiling

Tools

Docker
Kubernetes
CUDA

Job description

the company is pioneering the next generation of World Models, pushing the boundaries of generative AI, computer vision, and the simulation of reality. We are moving beyond 2D pixels to build models that natively understand the physics and geometry of our world. Our mission is to redefine how industries, from robotics and AR/VR to gaming and cinema, generate and interact with physically-grounded 3D environments.

We’re seeking a Research Engineer to own the backend that serves our world models. You will build and run the inference systems behind our product and API, turning frontier research models into fast, reliable, and scalable services. This is a hands‑on role for someone who has already owned production APIs in an AI system and can work closely with researchers to bring new models to users.

Responsibilities
  • Own the inference services and APIs that power our product, from request to generated output.
  • Optimize serving performance, reliability, and cost at scale. Maximize GPU utilization and minimize device memory footprint.
  • Apply latest techniques for improving inference performance.
  • Deploy research models as production services, including customer‑specific configurations.
  • Work with research and product colleagues to bring new models to users.
Key qualifications
  • 3+ years of software engineering in an AI, machine learning, or computer vision product environment.
  • Experience designing and owning production APIs.
  • Strong Python, and experience shipping containerized services on GPU inference platforms.
  • Hands‑on GPU inference performance work: profiling, memory, batching, and cold start.
  • Familiarity with 3D or computer vision data is a strong plus.

At the company, we are committed to creating a diverse and inclusive workplace. We welcome applications from people of all backgrounds, experiences, and perspectives. We are an equal opportunity employer and ensure all candidates are treated fairly throughout the recruitment process.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Inference Engineer
ML Inference Engineer

Spaitial Ltd. • Greater London

On-site
GBP 70,000 - 110,000
ML Inference Engineer
ML Inference Engineer

SpAItial AI • Greater London

On-site
GBP 70,000 - 110,000
ML Inference Engineer - Production APIs & GPU Performance
ML Inference Engineer - Production APIs & GPU Performance

United States Digital Space LLC • Greater London

On-site
GBP 80,000 - 120,000
GPU-Driven Inference Engineer for 3D World Models
GPU-Driven Inference Engineer for 3D World Models

Spaitial Ltd. • Greater London

On-site
GBP 70,000 - 110,000
Machine Learning & Cloud Infra Engineer
Machine Learning & Cloud Infra Engineer

SpAItial AI • Greater London

On-site
GBP 60,000 - 85,000
World Models Inference Engineer for 3D Vision APIs
World Models Inference Engineer for 3D Vision APIs

SpAItial AI • Greater London

On-site
GBP 70,000 - 110,000
Senior Software Engineer (Inference)
Senior Software Engineer (Inference)

AssemblyAI • Greater London

On-site
GBP 90,000 - 150,000
Home office stipend
Equity grant
Premium medical, dental, vision plans
+4
Senior Real-Time ML Inference Engineer
Senior Real-Time ML Inference Engineer

BITKRAFT Ventures • United Kingdom

On-site
GBP 140,000 - 200,000
Senior ML Infrastructure Engineer (Research Initiatives) - Systems Integrator
Senior ML Infrastructure Engineer (Research Initiatives) - Systems Integrator

Hamilton Barnes Associates Limited • United Kingdom

On-site
GBP 90,000 - 130,000
Significant stock option packages
Remote-first working setup
Fully paid travel and accommodation
+1
Research Engineer - 3D World Models
Research Engineer - 3D World Models

SpAItial AI • Greater London

On-site
GBP 40,000 - 60,000