Senior Systems Engineer, AI Inference Platform

Slope

San Francisco (CA)

On-site

USD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference optimization platform across hardware, compiler, and runtime contexts. You will design the OpenAI-hosted control plane and partner-side software, focusing on reliable long-running workflows, reproducible performance results, and secure handling of model and hardware data.

This cross-stack role collaborates with research, infrastructure, security, product, and partnerships to deliver scalable,

Qualifications

  • 8+ years of professional software engineering in large-scale distributed systems.
  • Strong programming in C++, Python, Go, or Rust.
  • Experience designing and operating backend systems and durable workflows for production workloads.
  • Strong understanding of distributed systems, Linux, networking, storage, containers, and cloud architectures.
  • Ability to lead complex technical initiatives as a senior individual contributor.

Responsibilities

  • Design, build, and operate durable APIs and control-plane services for multi-hour or multi-day optimization campaigns.
  • Build secure partner-side runner and grader software for compiling, executing, and benchmarking artifacts on accelerator hardware.
  • Integrate hardware profiles, toolchains, compilers, runtimes into repeatable optimization workflows.
  • Turn research prototypes into reliable product surfaces with contracts, debuggable failure modes, reproducible outputs.
  • Develop evaluation systems spanning latency, throughput, memory usage, utilization, and cost efficiency.
  • Build provenance and qualification workflows for safe review and deployment of optimized artifacts.
  • Collaborate with Research, Inference Engineering, Infrastructure, Security, Product, and Partnerships.
  • Drive architecture across cross-functional initiatives linking OpenAI systems with partner environments.

Skills

C++
Python
Go
Rust
Linux
Networking
Cloud

Tools

LLVM
MLIR
Triton
CUDA
ROCm

Job description

OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference optimization platform across hardware, compiler, and runtime contexts. You will design the OpenAI-hosted control plane and partner-side software, focusing on reliable long-running workflows, reproducible performance results, and secure handling of model and hardware data.

This cross-stack role collaborates with research, infrastructure, security, product, and partnerships to deliver scalable,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer for AI Inference & Optimization
Platform Engineer for AI Inference & Optimization

OpenAI • Seattle (WA)

On-site
USD 180,000 - 240,000
Software Engineer, AI Inference Infrastructure Platform
Software Engineer, AI Inference Infrastructure Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
AI Infra Systems Engineer — GPT Deployment & Optimization
AI Infra Systems Engineer — GPT Deployment & Optimization

OpenAI • United States

On-site
USD 180,000 - 260,000
Senior Systems Engineer, AI Inference Infra
Senior Systems Engineer, AI Inference Infra

United States Digital Space LLC • United States

Remote
USD 150,000 - 190,000
Senior Platform Engineer, OpenAI-compatible Inference API
Senior Platform Engineer, OpenAI-compatible Inference API

General Compute Inc. • San Francisco (CA)

On-site
USD 200,000 - 260,000
Senior AI Inference Infrastructure Engineer
Senior AI Inference Infrastructure Engineer

OpenAI • San Francisco (CA)

On-site
USD 293,000 - 445,000
Senior Platform Engineer – Inference API & Streaming
Senior Platform Engineer – Inference API & Streaming

General Compute Inc. • New York (NY)

On-site
USD 130,000 - 160,000
Host Systems Software Engineer: High-Performance AI Infra
Host Systems Software Engineer: High-Performance AI Infra

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Systems Generalist, GPT Infrastructure
Systems Generalist, GPT Infrastructure

OpenAI • Seattle (WA)

On-site
USD 293,000 - 445,000
Inference Platform Backend Engineer (Equity & Benefits)
Inference Platform Backend Engineer (Equity & Benefits)

Together • San Francisco (CA)

On-site
USD 160,000 - 250,000
Equity
Health insurance
Competitive compensation