Inference Engineer, Robotics

OpenAI

United States

Hybrid

USD 150,000 - 190,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Relocation assistance

Job summary

OpenAI's robotics team seeks a GPU Inference Engineer in San Francisco to advance model serving efficiency and inference performance. You will drive system optimizations, build serving infrastructure, and partner with researchers to scale multimodal workloads.

Ideal candidates combine deep expertise in inference optimization and low-level data movement with a collaborative approach to extending AI capabilities for real-world robotic applications. This role is hybrid with relocation assistance.

Qualifications

  • Deep expertise in model performance optimization, particularly at the inference layer.
  • Strong background in kernel-level systems, data movement, and low-level performance tuning.
  • Experience scaling AI systems that handle multimodal workloads.
  • Ability to navigate ambiguity, set technical direction, and drive complex initiatives.

Responsibilities

  • Perform engineering efforts focused on improving model serving, inference performance, and system efficiency.
  • Drive optimizations from a kernel and data movement perspective to improve system throughput and reliability.
  • Partner closely with research and product teams to ensure our models perform effectively at scale.
  • Design, build, and improve critical serving infrastructure to support Robotics growth and reliability needs.

Skills

Model performance optimization
Kernel-level systems
Data movement
Multimodal workloads
CUDA
Python
C/C++

Tools

CUDA
TensorRT

Job description

About the Team

Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples' lives.

About the Role

We're looking for a GPU Inference Engineer to contribute to improvements in model serving efficiency for our Robotics research. This is a high-impact role where you'll drive initiatives to optimize inference performance and scalability. You'll also be engaged in model design, to help assist our researchers in developing inference-friendly models. This role is critical to scaling the team's broader goals - it will directly enable leadership to focus on higher-leverage initiatives by building a stronger technical foundation. In this role you will:

  • Perform engineering efforts focused on improving model serving, inference performance, and system efficiency
  • Drive optimizations from a kernel and data movement perspective to improve system throughput and reliability
  • Partner closely with research and product teams to ensure our models perform effectively at scale
  • Design, build, and improve critical serving infrastructure to support Robotics growth and reliability needs
You might thrive in this role if you:
  • Have deep expertise in model performance optimization, particularly at the inference layer
  • Have a strong background in kernel-level systems, data movement, and low-level performance tuning
  • Are excited about scaling high-performing AI systems that serve real-world, multimodal workloads
  • Can navigate ambiguity, set technical direction, and drive complex initiatives to completion

This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees.

About OpenAI

OpenAI i

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Inference Engineer, Robotics
Inference Engineer, Robotics

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Relocation assistance
Hybrid work model (3 days in office)
GPU Inference Engineer, Robotics — Hybrid (Relocation)
GPU Inference Engineer, Robotics — Hybrid (Relocation)

OpenAI • United States

Hybrid
USD 150,000 - 190,000
Relocation assistance
Software Engineer, Model Inference
Software Engineer, Model Inference

OpenAI • San Francisco (CA)

On-site
USD 325,000 - 490,000
Senior Systems Engineer, AI Inference Platform
Senior Systems Engineer, AI Inference Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Software Engineer, GPT Infrastructure
Software Engineer, GPT Infrastructure

OpenAI • Seattle (WA)

On-site
USD 180,000 - 240,000
Software Engineer, GPT Infrastructure
Software Engineer, GPT Infrastructure

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Software Engineer, Inference - Performance Optimization
Software Engineer, Inference - Performance Optimization

OpenAI, Inc. • San Francisco (CA)

On-site
USD 295,000 - 555,000
Equity
INFERENCE ENGINEER
INFERENCE ENGINEER

MakerMaker.AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
AI Inference Engineer
AI Inference Engineer

Acceler8 Talent • San Francisco (CA)

On-site
USD 150,000 - 230,000
Software Engineer, Inference - Performance Optimization
Software Engineer, Inference - Performance Optimization

OpenAI • San Francisco (CA)

On-site
USD 295,000 - 555,000