On-Device ML Engineer — Mobile AI Optimizations

Unity Technologies

United States

Remote

USD 218,000 - 284,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health insurance
Stock options
Retirement plan
Generous vacation days
Parental leave support

Job summary

Unity Technologies is hiring a Senior Machine Learning Engineer for On-Device & Mobile AI to bring state-of-the-art multi-modal models to run fast and small on mobile and constrained hardware. You will own export, quantization, kernel tuning, and a shipped feature inside the engine at interactive frame rates.

You’ll work across NPUs, mobile GPUs, and desktop GPUs to shape latency, memory, and battery impact.

Qualifications

  • 5+ years in software/ML engineering with on-device/edge inference experience
  • Deployment of transformer- and/or diffusion-based models on mobile/desktop hardware
  • Hands-on with at least one major inference runtime (ONNX Runtime / ORT Web, CoreML, TFLite, ExecuTorch)
  • Low-level performance work: read a frame capture and kernel trace, profiling tools
  • Knowledge of quantization (INT4/INT8/FP16), weight sharing, pruning, distillation
  • Understanding of target hardware: mobile SoCs and GPUs (Apple/Qualcomm/ARM)
  • Strong Python; familiarity with browser-runtime languages (TypeScript/JavaScript, WGSL) is a plus
  • Working fluency with deployed models and architecture reasoning
  • Collaborative, reliable delivery and mentorship qualities

Responsibilities

  • Inference & On-Device Optimization: model export, graph transformation, operator fusion, memory-layout planning, hardware-specific tuning across NPUs and GPUs
  • Apply quantization, weight sharing, pruning, and distillation to meet latency/memory/power budgets and validate quality
  • Low-level performance work: write/tune WebGPU compute shaders and native kernels; profile with browser/platform tools
  • Build integration between ML runtime and game engine: real-time scheduling, memory pooling, zero-copy buffers, frame-budget management
  • Develop tooling and CI benchmarks for on-device performance and SKUs
  • Collaborate with research scientists to productionize CV and multi-modal architectures for on-device deployment
  • Provide feedback to research on hardware constraints and op-support gaps
  • Track breakthroughs in efficient inference and pragmatically apply movements that impact latency/memory/power
  • Share knowledge, review code, mentor junior engineers

Skills

On-device inference
WebGPU/WGSL
Model optimization
Python tooling
C++/Obj-C/Swift

Tools

Metal
Vulkan
CUDA
D3D12

Job description

Unity Technologies is hiring a Senior Machine Learning Engineer for On-Device & Mobile AI to bring state-of-the-art multi-modal models to run fast and small on mobile and constrained hardware. You will own export, quantization, kernel tuning, and a shipped feature inside the engine at interactive frame rates.

You’ll work across NPUs, mobile GPUs, and desktop GPUs to shape latency, memory, and battery impact.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

On-Device ML Engineer — Mobile AI & Optimization
On-Device ML Engineer — Mobile AI & Optimization

Unity • San Francisco (CA)

On-site
USD 180,000 - 240,000
Staff ML Engineer: On-Device AI for Real-Time Games
Staff ML Engineer: On-Device AI for Real-Time Games

Unity • Mountain View (CA)

On-site
USD 180,000 - 280,000
Senior On-Device ML Engineer – Mobile Inference Expert
Senior On-Device ML Engineer – Mobile Inference Expert

Unity • California (MO)

On-site
USD 180,000 - 240,000
Principal On-Device AI Engineer for Real-Time Game Inference
Principal On-Device AI Engineer for Real-Time Game Inference

Unity • Mountain View (CA)

On-site
USD 190,000 - 270,000
Principal On-Device AI Engineer – Real-Time Game Inference
Principal On-Device AI Engineer – Real-Time Game Inference

Unity Technologies • Mountain View (CA)

On-site
USD 278,000 - 418,000
Comprehensive health, life, and disability insurance
Employee stock ownership
Competitive retirement/pension plans
+1
ML Compiler Engineer - On-Device AI & ML
ML Compiler Engineer - On-Device AI & ML

Nutanix • Santa Clara (CA)

On-site
USD 151,000 - 227,000
Annual bonus
RSU grants
Benefits package
Staff Machine Learning Engineer
Staff Machine Learning Engineer

Unity Technologies • United States

Remote
USD 218,000 - 284,000
Health insurance
Stock options
Retirement plan
+2
Senior Real-Time ML Infrastructure Engineer
Senior Real-Time ML Infrastructure Engineer

3M HEALTHCARE • Bellevue (WA)

Remote
USD 183,000 - 249,000
Principal Machine Learning Engineer
Principal Machine Learning Engineer

Unity • Mountain View (CA)

On-site
USD 190,000 - 270,000
Staff On-Device ML Software Engineer
Staff On-Device ML Software Engineer

Google • Mountain View (CA)

On-site
USD 207,000 - 300,000