Tech TPM: Scale Model Deployment & Capacity

AI Chopping Block

San Francisco, Northern (CA, KY)

Hybrid

USD 190,000 - 230,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Relocation assistance
Hybrid work model

Job summary

OpenAI in San Francisco, CA is seeking a Technical Program Manager to lead the operating system for Chat capacity and model deployment. You will connect demand forecasting and capacity allocation with model readiness, rollout planning, and post-deployment learning.

You will own mainline model deployment across research, inference, and product teams, build scalable tooling, and define readiness gates, metrics, and clear communications for leaders.

Qualifications

  • Own cross-functional programs for Chat capacity forecasting, allocation, headroom planning, and constrained-capacity operations.
  • Build durable intake, prioritization, and decision mechanisms that connect product demand and model requirements to available serving capacity.
  • Partner with product, research, inference, fleet, and capacity teams to develop scenarios, surface tradeoffs, and drive timely allocation decisions.
  • Lead model deployment readiness and rollout planning, including serving-capacity allocation, launch sequencing, validation, and operational handoffs.
  • Establish clear readiness gates, risk reviews, rollback criteria, and escalation paths for model deployments.
  • Drive launch coordination through deployment and post-launch learning, turning recurring gaps and manual work into scalable tooling and operating practices.
  • Define and operationalize metrics for forecast accuracy, capacity utilization and headroom, deployment velocity, reliability, latency, quality, and user impact.
  • Create concise, decision-ready communications that make dependencies, risks, capacity constraints, and launch choices clear to technical and product leaders.

Responsibilities

  • Own cross-functional programs for Chat capacity forecasting, allocation, headroom planning, and constrained-capacity operations.
  • Build durable intake, prioritization, and decision mechanisms that connect product demand and model requirements to available serving capacity.
  • Partner with product, research, inference, fleet, and capacity teams to develop scenarios, surface tradeoffs, and drive timely allocation decisions.
  • Lead model deployment readiness and rollout planning, including serving-capacity allocation, launch sequencing, validation, and operational handoffs.
  • Establish clear readiness gates, risk reviews, rollback criteria, and escalation paths for model deployments.
  • Drive launch coordination through deployment and post-launch learning, turning recurring gaps and manual work into scalable tooling and operating practices.
  • Define and operationalize metrics for forecast accuracy, capacity utilization and headroom, deployment velocity, reliability, latency, quality, and user impact.
  • Create concise, decision-ready communications that make dependencies, risks, capacity constraints, and launch choices clear to technical and product leaders.

Skills

Technical program management
Program management
Distributed systems
Capacity planning
Model serving
Cross-functional collaboration
Communication

Job description

OpenAI in San Francisco, CA is seeking a Technical Program Manager to lead the operating system for Chat capacity and model deployment. You will connect demand forecasting and capacity allocation with model readiness, rollout planning, and post-deployment learning.

You will own mainline model deployment across research, inference, and product teams, build scalable tooling, and define readiness gates, metrics, and clear communications for leaders.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Strategic TPM: Model Deployment & Capacity
Strategic TPM: Model Deployment & Capacity

OpenAI • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Relocation assistance
Hybrid work model
Strategic TPM for Scale Model Deployments & Capacity
Strategic TPM for Scale Model Deployments & Capacity

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Relocation assistance
Technical Program Manager, Model Deployment & Capacity
Technical Program Manager, Model Deployment & Capacity

OpenAI • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Relocation assistance
Hybrid work model
Technical Program Manager, Model Deployment & Capacity
Technical Program Manager, Model Deployment & Capacity

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 190,000 - 230,000
Relocation assistance
Hybrid work model
Technical Program Manager, Model Deployment & Capacity
Technical Program Manager, Model Deployment & Capacity

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Relocation assistance
AI Safety & Deployment TPM
AI Safety & Deployment TPM

Slope • San Francisco (CA)

Hybrid
USD 160,000 - 230,000
Relocation assistance
Hybrid work model (3 days in SF office
Developer Experience TPM - Accelerate Dev Velocity
Developer Experience TPM - Accelerate Dev Velocity

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Relocation assistance
Hybrid work model (3 days in office)
GPU Compute Infrastructure TPM - Scale Large Clusters
GPU Compute Infrastructure TPM - Scale Large Clusters

OpenAI • California (MO)

Hybrid
USD 180,000 - 240,000
Relocation assistance
Hybrid work model
Gen AI Ops Planning TPM: Demand & Staffing Lead
Gen AI Ops Planning TPM: Demand & Staffing Lead

Scale • San Francisco (CA), New York (NY)

On-site
USD 151,000 - 189,000
Health insurance
Dental & vision coverage
Retirement benefits
+3
Lead AI Model Deployment & Experimentation
Lead AI Model Deployment & Experimentation

OpenAI • San Francisco (CA)

On-site
USD 210,000 - 300,000