Senior On-Device ML Engineer for Real-Time Multimodal AI

LE130 Unity Technologies SF

United States

Remote

USD 218,000 - 284,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Health and life insurance
Commuter subsidy
Employee stock ownership
Retirement plans
Generous vacation and personal days
Parental leave
Office snacks
Mental health programs
Employee Resource Groups
Training and development programs
Volunteer and donation matching

Job summary

Unity Technologies is seeking a Senior Machine Learning Engineer for On-Device & Mobile AI to optimize and deploy multi-modal models on mobile and constrained hardware. You will push federated optimization, export pipelines, and low-level GPU kernels to meet strict latency and memory budgets, collaborating with runtime and game engine teams.

The role emphasizes hands-on development, profiling, and cross-platform integration, with a focus on transformer- and diffusion-based models deployed

Qualifications

  • 5+ years in software/ML engineering.
  • Production deployment of transformer- and/or diffusion-based models on mobile/desktop hardware.
  • Hands-on experience with at least one major inference runtime (ONNX Runtime / ORT Web, CoreML, TFLite, ExecuTorch).
  • Low-level performance engineering with GPU/compute APIs and profiling tools.
  • Familiarity with model-optimization: quantization, pruning, distillation.

Responsibilities

  • Own the optimization pipeline for the models you ship: export, graph transformation, operator fusion, memory-layout planning, hardware tuning.
  • Write/tune WGSL compute shaders and native kernels; profile with browser/platform tools; eliminate bottlenecks.
  • Collaborate across runtimes and game engine integration: scheduling, memory pooling, zero-copy buffers, frame-budget management.
  • Partner with research to productionize CV and multi-modal architectures for on-device use.
  • Contribute to engineering standards, benchmarks, and CI to ensure performance.

Skills

Python
WebGPU/WGSL
C++/Objective-C/Swift
Profiling
On-device inference

Tools

ONNX Runtime Web
CoreML
TFLite
ExecuTorch

Job description

Unity Technologies is seeking a Senior Machine Learning Engineer for On-Device & Mobile AI to optimize and deploy multi-modal models on mobile and constrained hardware. You will push federated optimization, export pipelines, and low-level GPU kernels to meet strict latency and memory budgets, collaborating with runtime and game engine teams.

The role emphasizes hands-on development, profiling, and cross-platform integration, with a focus on transformer- and diffusion-based models deployed

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff ML Engineer: On-Device AI for Real-Time Games
Staff ML Engineer: On-Device AI for Real-Time Games

Unity • Mountain View (CA)

On-site
USD 180,000 - 280,000
On-Device ML Engineer — Mobile AI & Optimization
On-Device ML Engineer — Mobile AI & Optimization

Unity • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior On-Device ML Engineer – Mobile Inference Expert
Senior On-Device ML Engineer – Mobile Inference Expert

Unity • California (MO)

On-site
USD 180,000 - 240,000
On-Device ML Engineer - Fast, Lean AI for Games
On-Device ML Engineer - Fast, Lean AI for Games

Unity • California

Hybrid
USD 218,000 - 284,000
Health insurance
Life insurance
Stock ownership
+6
On-Device ML Engineer — Mobile AI Optimizations
On-Device ML Engineer — Mobile AI Optimizations

Unity Technologies • United States

Remote
USD 218,000 - 284,000
Health insurance
Stock options
Retirement plan
+2
Principal On-Device AI Engineer for Real-Time Game Inference
Principal On-Device AI Engineer for Real-Time Game Inference

Unity • Mountain View (CA)

On-site
USD 190,000 - 270,000
Lead On-Device AI Inference Engineer for Next-Gen Games
Lead On-Device AI Inference Engineer for Next-Gen Games

LE130 Unity Technologies SF • Mountain View (CA)

On-site
USD 260,000 - 380,000
Senior ML Engineer: Real-Time AI Agentic Systems
Senior ML Engineer: Real-Time AI Agentic Systems

Unity Technologies SF • Mountain View (CA)

On-site
USD 210,000 - 273,000
Health insurance
Stock options
Pension plan
+6
Senior Real-Time ML Infrastructure Engineer
Senior Real-Time ML Infrastructure Engineer

3M HEALTHCARE • Bellevue (WA)

Remote
USD 183,700 - 248,600
Senior AI Engineer for Real-Time Agentic Systems
Senior AI Engineer for Real-Time Agentic Systems

LE130 Unity Technologies SF • Mountain View (CA)

On-site
USD 167,000 - 245,000
Health insurance
Commuter benefits
Equity and annual bonus
+6