GPU Inference Engineer, Robotics — Hybrid (Relocation)

OpenAI

United States

Hybrid

USD 150,000 - 190,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Relocation assistance

Job summary

OpenAI's robotics team seeks a GPU Inference Engineer in San Francisco to advance model serving efficiency and inference performance. You will drive system optimizations, build serving infrastructure, and partner with researchers to scale multimodal workloads.

Ideal candidates combine deep expertise in inference optimization and low-level data movement with a collaborative approach to extending AI capabilities for real-world robotic applications. This role is hybrid with relocation assistance.

Qualifications

  • Deep expertise in model performance optimization, particularly at the inference layer.
  • Strong background in kernel-level systems, data movement, and low-level performance tuning.
  • Experience scaling AI systems that handle multimodal workloads.
  • Ability to navigate ambiguity, set technical direction, and drive complex initiatives.

Responsibilities

  • Perform engineering efforts focused on improving model serving, inference performance, and system efficiency.
  • Drive optimizations from a kernel and data movement perspective to improve system throughput and reliability.
  • Partner closely with research and product teams to ensure our models perform effectively at scale.
  • Design, build, and improve critical serving infrastructure to support Robotics growth and reliability needs.

Skills

Model performance optimization
Kernel-level systems
Data movement
Multimodal workloads
CUDA
Python
C/C++

Tools

CUDA
TensorRT

Job description

OpenAI's robotics team seeks a GPU Inference Engineer in San Francisco to advance model serving efficiency and inference performance. You will drive system optimizations, build serving infrastructure, and partner with researchers to scale multimodal workloads.

Ideal candidates combine deep expertise in inference optimization and low-level data movement with a collaborative approach to extending AI capabilities for real-world robotic applications. This role is hybrid with relocation assistance.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Inference Engineer, Robotics
Inference Engineer, Robotics

OpenAI • United States

Hybrid
USD 150,000 - 190,000
Relocation assistance
Inference Engineer, Robotics
Inference Engineer, Robotics

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Relocation assistance
Hybrid work model (3 days in office)
Staff Engineer, GPU AI Inference & RL Infrastructure
Staff Engineer, GPU AI Inference & RL Infrastructure

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive medical, dental, and vision insurance
Fully paid parental leave
+2
Hybrid Performance Modeling Engineer - Relocation Available
Hybrid Performance Modeling Engineer - Relocation Available

OpenAI • United States

Hybrid
USD 90,000 - 130,000
Hybrid work model (3 days in office)
Relocation assistance
Real-Time GPU Optimization Engineer - Inference
Real-Time GPU Optimization Engineer - Inference

techire ai • San Francisco (CA)

On-site
USD 230,000 - 300,000
Senior AI Kernel Engineer — Remote/Hybrid Inference
Senior AI Kernel Engineer — Remote/Hybrid Inference

Modular • United States

Hybrid
USD 198,000 - 286,000
Amazing Team
World-class Benefits
Competitive Compensation
+1
Senior GPU Inference Engine Engineer
Senior GPU Inference Engine Engineer

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Flexible working hours
Daily lunch and dinner provided; unlimited snacks and beverages
Health check-up support and top-tier equipment/hardware support
+2
AI Hardware Co-Design Engineer — Hybrid (SF)
AI Hardware Co-Design Engineer — Hybrid (SF)

OpenAI • United States

Hybrid
USD 150,000 - 190,000
Relocation assistance
GPU Infra Engineer for Scalable AI Platform
GPU Infra Engineer for Scalable AI Platform

Kindredventures • San Francisco (CA)

On-site
USD 180,000 - 250,000
Relocation to San Francisco
Health, dental, and vision insurance (
Regular team events
+1
Senior AI Inference Systems Engineer (GPU, Open Source)
Senior AI Inference Systems Engineer (GPU, Open Source)

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 224,000 - 431,000
Equity
Employee benefits