Accelerator Systems Software Engineer for AI Compute

OpenAI

California (MO)

On-site

USD 180,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenAI’s Accelerators team is seeking engineers to evaluate and bring up new compute platforms that can support large-scale AI training and inference. You will prototype system software on accelerators and drive performance across AI workloads in a tightly integrated stack.

You’ll collaborate across hardware and software, developing kernels, sharding strategies, and scaling methods to enable frontier AI workloads at scale.

Qualifications

  • 3+ years of experience working on AI infrastructure, including kernels, systems, or hardware-software co-design.
  • Hands-on experience with accelerator platforms for AI at data center scale.
  • Strong understanding of kernels, sharding, runtime systems, or distributed scaling techniques.
  • Familiarity with optimizing LLMs, CNNs, or recommender models for hardware efficiency.
  • Experience with performance modeling, system debugging, and software stack adaptation for novel architectures.
  • Exposure to mobile accelerators is welcome, but data center-scale AI hardware experience is preferred.

Responsibilities

  • Prototype and enable OpenAI's AI software stack on new, exploratory accelerator platforms.
  • Optimize large-scale model performance for diverse hardware environments.
  • Develop kernels, sharding mechanisms, and system scaling strategies tailored to emerging accelerators.
  • Collaborate on optimizations at the model code level (e.g. PyTorch) and below to enhance performance on non-traditional hardware.
  • Perform system-level performance modeling, debug bottlenecks, and drive end-to-end optimization.
  • Work with hardware teams and vendors to evaluate alternatives to existing platforms and adapt the software stack to their architectures.
  • Contribute to runtime improvements, compute/communication overlapping, and scaling efforts for frontier AI workloads.

Skills

AI infrastructure
Kernels
Sharding
Runtime systems
Distributed scaling
PyTorch
Performance modeling

Tools

TPUs
Custom silicon

Job description

OpenAI’s Accelerators team is seeking engineers to evaluate and bring up new compute platforms that can support large-scale AI training and inference. You will prototype system software on accelerators and drive performance across AI workloads in a tightly integrated stack.

You’ll collaborate across hardware and software, developing kernels, sharding strategies, and scaling methods to enable frontier AI workloads at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Accelerators
Software Engineer, Accelerators

OpenAI • California (MO)

On-site
USD 180,000 - 300,000
Software Engineer, Accelerators
Software Engineer, Accelerators

Slope • San Francisco (CA)

On-site
USD 120,000 - 180,000
AI Transformation & Accelerator Systems Architect
AI Transformation & Accelerator Systems Architect

Google • Sunnyvale (CA)

On-site
USD 262,000 - 364,000
AI Accelerator Runtime Engineer (Low-Level Systems)
AI Accelerator Runtime Engineer (Low-Level Systems)

OpenAI • San Francisco (CA)

On-site
USD 266,000 - 445,000
Platform Engineer for AI Inference & Optimization
Platform Engineer for AI Inference & Optimization

OpenAI • Seattle (WA)

On-site
USD 180,000 - 240,000
Software Engineer, AI Inference Infrastructure Platform
Software Engineer, AI Inference Infrastructure Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
AI Accelerator Runtime Systems Engineer
AI Accelerator Runtime Systems Engineer

Triwill Group • United States

On-site
USD 180,000 - 240,000
AI Accelerator Systems Program Manager
AI Accelerator Systems Program Manager

OpenAI • San Francisco (CA)

Hybrid
USD 302,000 - 445,000
Relocation assistance
Hybrid work model
Hardware-Focused Software Engineer for AI Systems
Hardware-Focused Software Engineer for AI Systems

OpenAI • California (MO)

Hybrid
USD 180,000 - 280,000
Relocation assistance
AI Accelerator Systems Software TPM (Hybrid)
AI Accelerator Systems Software TPM (Hybrid)

Triwill Group • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Relocation assistance