AI Accelerator Runtime Engineer (Low-Level Systems)

OpenAI

San Francisco (CA)

On-site

USD 266,000 - 445,000

Full time

9 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenAI is seeking a highly technical engineer to build the low-level device runtime for its custom AI accelerator and to optimize kernel launches, memory management, and synchronization across hardware and software boundaries.

You will work with compiler, kernel, architecture, verification, and silicon teams, using cycle-accurate simulators to validate behavior before silicon becomes available and after deployment to ensure performance and correctness.

Qualifications

  • Strong low-level systems programming experience in C, C++, Rust, or comparable environments.
  • Built runtimes, drivers, firmware, operating-system components, accelerator software, or adjacent systems infrastructure.
  • Understand concurrency, synchronization, asynchronous execution, queues, events, and memory-ordering semantics.

Responsibilities

  • Design and implement the low-level device runtime for OpenAI custom silicon.
  • Build kernel-launch scheduling, command submission, queueing, dependency tracking, and completion handling.
  • Manage device memory spaces, allocation, virtual-to-physical mappings, data movement, and lifetime across concurrent workloads.
  • Implement synchronization primitives, events, barriers, streams, and ordering guarantees that are correct and efficient.
  • Define clean interfaces between the runtime, drivers, firmware, compiler-generated code, kernels, and higher-level execution systems.
  • Use event-based, cycle-accurate simulators to develop, validate, debug, and performance-tune runtime behavior before and after silicon availability.
  • Diagnose concurrency, memory-ordering, deadlock, race, correctness, and performance issues across software and hardware boundaries.
  • Build tests, tracing, profiling, observability, and reproducible workloads for runtime correctness and performance.
  • Partner with architecture and silicon teams to turn workload and simulator insights into hardware-software interface improvements.

Skills

Low-level systems
C/C++
Rust
Concurrency

Tools

Event-based simulators

Job description

OpenAI is seeking a highly technical engineer to build the low-level device runtime for its custom AI accelerator and to optimize kernel launches, memory management, and synchronization across hardware and software boundaries.

You will work with compiler, kernel, architecture, verification, and silicon teams, using cycle-accurate simulators to validate behavior before silicon becomes available and after deployment to ensure performance and correctness.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Low-Level AI Accelerator Runtime Engineer
Low-Level AI Accelerator Runtime Engineer

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
AI Accelerator Runtime Systems Engineer
AI Accelerator Runtime Systems Engineer

Triwill Group • United States

On-site
USD 180,000 - 240,000
Software Engineer, AI accelerator Runtime
Software Engineer, AI accelerator Runtime

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Software Engineer, AI accelerator Runtime
Software Engineer, AI accelerator Runtime

Triwill Group • United States

On-site
USD 180,000 - 240,000
Software Engineer, AI accelerator Runtime
Software Engineer, AI accelerator Runtime

OpenAI • San Francisco (CA)

On-site
USD 266,000 - 445,000
AI Hardware Systems Engineer
AI Hardware Systems Engineer

Triwill Group • United States

On-site
USD 180,000 - 230,000
AI Runtime Engineer
AI Runtime Engineer

Insider, Inc. • United States

On-site
USD 100,000 - 140,000
Software Engineer, Kernel Performance & AI Tooling
Software Engineer, Kernel Performance & AI Tooling

OpenAI • Los Angeles (CA)

On-site
USD 100,000 - 130,000
Senior AI Kernel & Performance Engineer | Equity
Senior AI Kernel & Performance Engineer | Equity

Meta • Menlo Park (CA)

On-site
USD 154,000 - 217,000
Compiler Runtime Engineer
Compiler Runtime Engineer

Oho Group • San Francisco (CA)

On-site
USD 150,000 - 210,000