Developer Productivity Engineer — Inference Platform

Neura Market

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenAI is hiring a Developer Productivity Engineer to support OpenAI’s Inference Runtime teams. The role focuses on scaling engineering systems, safeguards, and developer workflows to enable rapid, reliable model deployment and inference at scale.

You’ll work on tooling for deploy gates, release validation, and production readiness across a demanding inference stack. You’ll collaborate with cross‑functional teams to improve automation, observability, and rollout safety, ensuring images released

Qualifications

  • Experience with CI/CD systems and large-scale build/validation
  • Strong focus on developer productivity and reliability
  • Ability to improve tooling and automation across complex stacks

Responsibilities

  • Improve deploy gate validation tooling and infrastructure
  • Enhance release processes and branching across the inference stack
  • Strengthen canary, async, and large-scale validation workflows
  • Harden CI, testing, and validation to make failures actionable
  • Reduce flaky failures from infrastructure, GPU scheduling, or test environments
  • Build automation for triage, ownership detection, debugging, and escalation
  • Collaborate with inference teams to improve release quality and rollout safety
  • Reduce developer friction in testing and release workflows

Skills

CI/CD systems
Testing infrastructure
Release tooling
Developer productivity
Large-scale build and validation

Tools

Python
C++

Job description

OpenAI is hiring a Developer Productivity Engineer to support OpenAI’s Inference Runtime teams. The role focuses on scaling engineering systems, safeguards, and developer workflows to enable rapid, reliable model deployment and inference at scale.

You’ll work on tooling for deploy gates, release validation, and production readiness across a demanding inference stack. You’ll collaborate with cross‑functional teams to improve automation, observability, and rollout safety, ensuring images released

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Productivity - Inference Runtime
Software Engineer, Productivity - Inference Runtime

OpenAI • Los Angeles (CA)

On-site
USD 230,000 - 385,000
Inference Runtime Developer Productivity Engineer
Inference Runtime Developer Productivity Engineer

OpenAI • Los Angeles (CA)

On-site
USD 230,000 - 385,000
Software Engineer, Productivity - Inference Runtime
Software Engineer, Productivity - Inference Runtime

Neura Market • San Francisco (CA)

On-site
USD 180,000 - 240,000
Developer Productivity Engineer, Model Performance
Developer Productivity Engineer, Model Performance

OpenAI • San Francisco (CA)

On-site
USD 230,000 - 385,000
Software Engineer, Developer Experience & Productivity
Software Engineer, Developer Experience & Productivity

OpenAI • San Francisco (CA)

On-site
USD 230,000 - 385,000
Senior AI Inference Infrastructure Engineer
Senior AI Inference Infrastructure Engineer

OpenAI • San Francisco (CA)

On-site
USD 293,000 - 445,000
Senior Systems Engineer, AI Inference Platform
Senior Systems Engineer, AI Inference Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Software Engineer, AI Inference Infrastructure Platform
Software Engineer, AI Inference Infrastructure Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Platform Engineer for AI Inference & Optimization
Platform Engineer for AI Inference & Optimization

OpenAI • Seattle (WA)

On-site
USD 180,000 - 240,000
AI Inference Platform Engineer — Production ML Tools
AI Inference Platform Engineer — Production ML Tools

Baseten • United States

Remote
USD 140,000 - 210,000
Equity
Medical insurance
Dental & Vision
+4