Senior ML Infra Engineer — Real-Time Inference

Unity Technologies SF

Olympia (WA)

On-site

USD 166,000 - 273,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Comprehensive health insurance
Commute subsidy
Employee stock ownership
Generous vacation and personal days
Mental Health programs
Volunteering and donation matching

Job summary

Unity Technologies SF is seeking a Senior Machine Learning Engineer to design and evolve Unity Vector’s online inference platform. The role focuses on production infrastructure for low-latency inference, with emphasis on optimization, observability, and safe experimentation.

You will design and operate large-scale online inference infrastructure, build distributed training support, and drive end-to-end serving lifecycle management.

Qualifications

  • Experience building production-grade online ML inference systems.
  • Experience with model serving frameworks and distributed serving platforms.
  • Strong background in distributed systems and production observability.
  • Proficiency in Python for high-scale services.
  • Experience with PyTorch and modern deployment workflows.

Responsibilities

  • Design and operate large-scale online inference infrastructure for production ML models.
  • Build infrastructure for distributed training workflows.
  • Integrate ML pipelines with workflow orchestration systems.
  • Optimize model performance and resource utilization.
  • Enhance ML system observability and reliability.
  • Collaborate to accelerate model iteration while maintaining safety.
  • Improve deployment automation and reproducibility of serving workflows.
  • Lead architectural improvements for robustness and cost efficiency.

Skills

Python programming
Distributed systems
Kubernetes
ML inference
Low-latency optimization
Observability
Canary / A/B testing
Canary testing
Production ML systems

Tools

PyTorch
NVIDIA Triton Inference Server
TorchServe
Ray Serve
TensorFlow Serving
Kubernetes
GKE
Ray
Flyte
Airflow
Ray Data
Ray Train
Python

Job description

Unity Technologies SF is seeking a Senior Machine Learning Engineer to design and evolve Unity Vector’s online inference platform. The role focuses on production infrastructure for low-latency inference, with emphasis on optimization, observability, and safe experimentation.

You will design and operate large-scale online inference infrastructure, build distributed training support, and drive end-to-end serving lifecycle management.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior ML Infra Engineer: Production Inference
Senior ML Infra Engineer: Production Inference

Unity • United States

Remote
USD 187,000 - 260,000
Health insurance
Stock options
Retirement plans
+1
Senior ML Infra Engineer — Production Inference Architect
Senior ML Infra Engineer — Production Inference Architect

LE130 Unity Technologies SF • United States

Remote
USD 210,000 - 273,000
Health insurance
Stock options
Retirement plans
+2
Senior ML Infrastructure Engineer - Online Inference
Senior ML Infrastructure Engineer - Online Inference

Unity Enterprise • Northern (KY)

Hybrid
USD 210,000 - 273,000
Comprehensive health insurance
Employee stock ownership
Generous vacation and personal days
+1
Senior Real-Time ML Infrastructure Engineer
Senior Real-Time ML Infrastructure Engineer

3M HEALTHCARE • Bellevue (WA)

Remote
USD 183,700 - 248,600
Remote Senior ML Infra Engineer — Real-Time Model Serving
Remote Senior ML Infra Engineer — Real-Time Model Serving

3M HEALTHCARE • Mountain View (CA)

On-site
USD 183,700 - 248,600
Health insurance
Stock options
Retirement plan
+1
Remote Senior ML Infra Engineer — Real-Time Model Serving
Remote Senior ML Infra Engineer — Real-Time Model Serving

3M HEALTHCARE • Mountain View (CA)

On-site
USD 183,700 - 248,600
Health insurance
Stock options
Retirement plan
+1
Senior On-Device ML Engineer – Mobile Inference Expert
Senior On-Device ML Engineer – Mobile Inference Expert

Unity • California (MO)

On-site
USD 180,000 - 240,000
Senior ML Engineer: Real-Time AI Agentic Systems
Senior ML Engineer: Real-Time AI Agentic Systems

Unity Technologies SF • Mountain View (CA)

On-site
USD 210,000 - 273,000
Health insurance
Stock options
Pension plan
+6
Senior ML Engineer, Data Infrastructure & Pipelines
Senior ML Engineer, Data Infrastructure & Pipelines

Unity Enterprise • Mountain View (CA)

Hybrid
USD 200,000 - 261,000
Comprehensive health insurance
Life and disability insurance
Employee stock ownership
+3
Staff ML Engineer: On-Device AI for Real-Time Games
Staff ML Engineer: On-Device AI for Real-Time Games

Unity • Mountain View (CA)

On-site
USD 180,000 - 280,000