On-Device AI Inference OS Engineer

OpenAI

San Francisco (CA)

On-site

USD 230,000 - 385,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Medical benefits
Dental benefits
Vision benefits
401(k) matching
Parental leave
Paid PTO
Relocation help
Learning stipend
Meal programs

Job summary

OpenAI in San Francisco is seeking an Operating Systems Engineer focused on on-device inference to design, develop, and ship the OS stack for reliable, energy-efficient AI on consumer devices. You will work across OS services, inference runtimes, memory management, and scheduling, partnering with research to adapt models to device constraints and push quality while honoring power and latency targets.

This role combines hardware, software, and AI, and offers collaboration across research,

Qualifications

  • Substantial hands-on experience designing, developing, and debugging OS components and services.
  • Proficiency in C++ for systems development with concurrency and memory management.
  • Experience optimizing inference runtimes in resource-constrained environments.
  • Strong scheduling, memory management, and shared resource understanding.
  • Ability to diagnose complex system behavior with measure-based improvements.
  • Ability to translate research and product needs into system requirements.

Responsibilities

  • Build the inference platform with OS services, interfaces, and lifecycle management.
  • Fit models to device constraints through quantization and memory optimization.
  • Coordinate system resources with scheduling policies for latency and power.
  • Advance performance and power management with adaptive execution strategies.
  • Debug across the stack using tracing and profiling for reliability.
  • Measure improvements with diagnostics and automated tests on devices.
  • Bring capabilities to production in collaboration with multiple teams.

Skills

C++
OS design
Inference runtimes
Scheduling
Performance
Cross-functional
Rust

Job description

OpenAI in San Francisco is seeking an Operating Systems Engineer focused on on-device inference to design, develop, and ship the OS stack for reliable, energy-efficient AI on consumer devices. You will work across OS services, inference runtimes, memory management, and scheduling, partnering with research to adapt models to device constraints and push quality while honoring power and latency targets.

This role combines hardware, software, and AI, and offers collaboration across research,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

On-Device AI OS Engineer: Inference & Performance
On-Device AI OS Engineer: Inference & Performance

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Operating Systems Engineer, On-Device Inference | Consumer Devices
Operating Systems Engineer, On-Device Inference | Consumer Devices

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior Systems Engineer, AI Inference Platform
Senior Systems Engineer, AI Inference Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Platform Engineer for AI Inference & Optimization
Platform Engineer for AI Inference & Optimization

OpenAI • Seattle (WA)

On-site
USD 180,000 - 240,000
On-Device AI Systems Engineer
On-Device AI Systems Engineer

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 150,000 - 278,000
Medical and dental coverage
Retirement benefits
Discounted products and services
+3
Software Engineer, AI Inference Infrastructure Platform
Software Engineer, AI Inference Infrastructure Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Inference Runtime Engineer - On-Device & Cloud AI, Flexible WFH
Inference Runtime Engineer - On-Device & Cloud AI, Flexible WFH

EngRadar • New York (NY)

On-site
USD 150,000 - 230,000
Equity grants
Medical plan
Vision plan
+5
AI Silicon Bringup & Systems Engineer
AI Silicon Bringup & Systems Engineer

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 230,000 - 300,000
Linux Kernel Systems Engineer for AI Devices
Linux Kernel Systems Engineer for AI Devices

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
OS Architect for Consumer Devices
OS Architect for Consumer Devices

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000