Data Scientist - Inference & GPU Capacity Optimization

OpenAI

United States

On-site

USD 150,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenAI is seeking a Data Scientist to partner with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. You will build forecasting models, analyze workloads, and design experiments to drive infrastructure investments and performance efficiency while communicating insights to executives.

Ideal candidates bring 5+ years in infrastructure data science, strong Python/SQL skills, and a track record of translating complex data

Qualifications

  • MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience).
  • 5+ years of experience working in the infrastructure data science space.
  • Strong expertise in Python and SQL.
  • Experience building forecasting, optimization, or predictive models.
  • Strong understanding of experimentation, statistical inference, and causal analysis.
  • Experience communicating analytical insights to executive stakeholders.

Responsibilities

  • Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency.
  • Develop forecasting models for inference demand across products, regions, and model families.
  • Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities.
  • Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies.
  • Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs.
  • Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions.
  • Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps.
  • Communicate technical findings clearly to both engineering teams and executive leadership.

Skills

Forecasting
Experimental design
Communication with executives

Education

MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline

Tools

Python
SQL

Job description

OpenAI is seeking a Data Scientist to partner with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. You will build forecasting models, analyze workloads, and design experiments to drive infrastructure investments and performance efficiency while communicating insights to executives.

Ideal candidates bring 5+ years in infrastructure data science, strong Python/SQL skills, and a track record of translating complex data

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist, Inference Capacity Optimization
Data Scientist, Inference Capacity Optimization

OpenAI • United States

On-site
USD 150,000 - 210,000
Data Scientist — Inference Capacity & GPU Optimization
Data Scientist — Inference Capacity & GPU Optimization

United States Digital Space LLC • United States

Remote
USD 140,000 - 230,000
Remote Senior AI Inference Optimization Engineer
Remote Senior AI Inference Optimization Engineer

DigitalOcean • San Francisco (CA)

On-site
USD 191,000 - 239,000
Equity compensation
Remote work
GPU Infra Engineer for Scalable AI Compute
GPU Infra Engineer for Scalable AI Compute

OpenAI • United States

On-site
USD 180,000 - 240,000
Lead, AI Inference & GPU Strategy
Lead, AI Inference & GPU Strategy

DigitalOcean • Seattle (WA)

Hybrid
USD 218,000 - 273,000
Senior AI Inference Infrastructure Engineer
Senior AI Inference Infrastructure Engineer

OpenAI • San Francisco (CA)

On-site
USD 293,000 - 445,000
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 280,000 - 420,000
Equity options
Health, vision, dental benefits
Unlimited PTO
+2
GPU Inference Performance Engineer — Equity & Optimization
GPU Inference Performance Engineer — Equity & Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Staff Engineer, GPU AI Inference & RL Infrastructure
Staff Engineer, GPU AI Inference & RL Infrastructure

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive medical, dental, and vision insurance
Fully paid parental leave
+2
Senior AI Infrastructure Engineer — Scale GPU Clusters
Senior AI Infrastructure Engineer — Scale GPU Clusters

AI Breaking Wire • San Francisco (CA)

On-site
USD 280,000 - 400,000
Equity
Medical, dental, and vision benefits
Unlimited PTO
+2