GPU Inference Capacity Optimization Scientist

openai

San Francisco (CA)

On-site

USD 293,000 - 325,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

OpenAI is seeking a Data Scientist to partner with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. You will transform complex operational data into actionable insights that influence how we allocate and scale one of the world's largest AI compute environments.

You will build models to profile GPU utilization, develop forecasting for inference demand, and design experiments to evaluate scheduling policies and

Qualifications

  • MS or PhD in statistics, CS, OR, applied math, economics, or related quantitative field.
  • 5+ years of experience in the infrastructure data science space.
  • Strong Python and SQL expertise.
  • Experience in forecasting, optimization, or predictive models.
  • Strong understanding of experimentation, statistical inference, and causal analysis.
  • Experience communicating insights to executive stakeholders.

Responsibilities

  • Build statistical and ML models to profile GPU utilization, latency, throughput, and fleet efficiency.
  • Develop forecasting models for inference demand across products, regions, and model families.
  • Analyze workloads to identify latency bottlenecks and capacity constraints.
  • Inform infrastructure planning and long-term GPU investment strategies with Capacity Systems Engineering.
  • Design experiments and simulations to evaluate scheduling policies and tradeoffs.
  • Build dashboards and metrics for data-driven capacity decisions by leadership.
  • Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with roadmaps.
  • Communicate technical findings clearly to engineering teams and executives.

Skills

Python
SQL
Forecasting
Statistical inference
Causal analysis
Executive communication

Education

MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline

Tools

Python
SQL

Job description

OpenAI is seeking a Data Scientist to partner with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. You will transform complex operational data into actionable insights that influence how we allocate and scale one of the world's largest AI compute environments.

You will build models to profile GPU utilization, develop forecasting for inference demand, and design experiments to evaluate scheduling policies and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Scientist, Inference Capacity Optimization
Data Scientist, Inference Capacity Optimization

openai • San Francisco (CA)

On-site
USD 293,000 - 325,000
Senior GPU Capacity & Optimization Architect
Senior GPU Capacity & Optimization Architect

Epoch Biodesign • United States

On-site
USD 160,000 - 195,000
Competitive compensation
Paid time off
Comprehensive health insurance
+3
Strategic Finance Lead for GPU Compute & Capacity Planning
Strategic Finance Lead for GPU Compute & Capacity Planning

Perplexity AI • United States

On-site
USD 140,000 - 210,000
Senior GPU Capacity & Optimization Architect
Senior GPU Capacity & Optimization Architect

Crusoe • San Francisco (CA)

On-site
USD 160,000 - 195,000
Competitive compensation and equity packages
Comprehensive health, dental & vision insurance
401(k) Retirement plan with company match
+2
Senior AI Infrastructure Engineer — Scale GPU Clusters
Senior AI Infrastructure Engineer — Scale GPU Clusters

AI Breaking Wire • San Francisco (CA)

On-site
USD 280,000 - 400,000
Equity
Medical, dental, and vision benefits
Unlimited PTO
+2
Senior AI Inference Optimization Engineer
Senior AI Inference Optimization Engineer

DigitalOcean • Seattle (WA)

On-site
USD 191,000 - 239,000
Lead Capacity Strategy & Operations (GPU Demand)
Lead Capacity Strategy & Operations (GPU Demand)

Baseten • San Francisco (CA)

On-site
USD 190,000 - 235,000
Equity
Health insurance
Flexible PTO
+3
Inference Performance Engineer: Cost & Capacity Modeling
Inference Performance Engineer: Cost & Capacity Modeling

Visa Hunt • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Head of Capacity Strategy & AI Compute Ops
Head of Capacity Strategy & AI Compute Ops

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Equity
Full medical coverage
Flexible PTO
+3
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 280,000 - 420,000
Equity options
Health, vision, dental benefits
Unlimited PTO
+2