Technical Program Manager (Inference)

CoreWeave

York and North Yorkshire

On-site

GBP 80,000 - 110,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

CoreWeave is seeking an experienced Technical Program Manager focused on inference to lead cross-functional programs across the AI/ML Platform Services organization. You will partner with engineering, product, research, and GTM teams to deliver scalable, reliable inference platforms for researchers, engineers, and enterprise customers.

You will drive onboarding, launch readiness, and runtime optimization for serverless and dedicated inference, building dashboards and gates to ensure high

Qualifications

  • Understanding of customer onboarding for technical products, especially where platform capabilities, infrastructure readiness, and support processes must align for launch.
  • Bachelor’s degree in a technical field or equivalent practical experience.
  • Proven experience driving large-scale infrastructure or platform programs from concept to production in complex, cross-functional environments.
  • Excellent written and verbal communication skills, with the ability to align engineering, product, infrastructure, and customer-facing stakeholders around shared goals.
  • Experience with inference-serving systems, model onboarding workflows, rollout strategies, and observability tooling.
  • Familiarity with launch readiness, supportability, incident follow-through, and release validation for production infrastructure or platform services.
  • 8+ years of technical program management experience in distributed systems, cloud infrastructure, or AI/ML platform engineering.
  • Demonstrated success driving measurable improvements in reliability, performance, operational readiness, or customer delivery.
  • Experience operating in high-growth environments where roadmap execution, reliability expectations, and customer commitments must be managed in parallel.
  • You’re effective at creating clarity and momentum across ambiguous, fast-moving multi-team initiatives.
  • You love driving execution for complex, customer-facing AI infrastructure and platform programs.
  • You’re curious about how large-scale inference systems evolve across runtime performance, operational excellence, and customer onboarding.
  • You enjoy turning technically complex platform work into predictable execution and successful launches.

Responsibilities

  • Own delivery and execution across CoreWeave’s AI/ML Platform Services organization.
  • Partner with Product, Engineering, Research, Infrastructure, and Go-to-Market teams to deliver scalable, reliable platforms that support the full AI lifecycle.
  • Drive alignment and execution across highly technical, cross-functional teams to ensure the successful delivery of customer-facing infrastructure and platform capabilities used by researchers, engineers, and enterprise customers.
  • Lead complex, cross-functional programs spanning inference platform delivery, customer onboarding, launch readiness, and runtime optimization.
  • Build and operate highly scalable, reliable production inference services for serverless and dedicated inference use cases.
  • Create strong success metrics, dashboards, launch gates, and review cadences to measure service reliability, onboarding readiness, efficiency, and quality across the inference stack.
  • Establish repeatable processes for release validation, performance regression tracking, launch management, and postmortem follow-through.
  • Help unify operational processes, support mechanisms, and execution visibility across inference deployment models and customer onboarding paths.
  • Create strong communication channels between Engineering, Product, Infrastructure, and Go-to-Market teams to align priorities and deliver predictable, high-impact outcomes.

Skills

Distributed inference
GPU compute
Cloud-native
Performance optimization
Customer onboarding
Cross-functional leadership
Program management
Observability tooling
Reliability engineering
Communication

Education

Bachelor’s degree in a technical field

Job description

Responsibilities
  • The AI/ML TPM team owns delivery and execution across CoreWeave’s AI/ML Platform Services organization
  • The team partners closely with Product, Engineering, Research, Infrastructure, and Go-to-Market teams to deliver scalable, reliable, and high-performance platforms that support the full AI lifecycle
  • AI/ML TPMs drive alignment and execution across highly technical, cross-functional teams to ensure the successful delivery of customer-facing infrastructure and platform capabilities used by researchers, engineers, and enterprise customers
  • As a Technical Program Manager focused on inference, you will lead complex, cross-functional programs spanning inference platform delivery, customer onboarding, launch readiness, and runtime optimization
  • The Inference team is responsible for building and operating highly scalable, reliable production inference services for both serverless and dedicated inference use cases
  • The work spans platform reliability, operational excellence, customer onboarding, release validation, runtime performance, and the launch of new inference capabilities that help customers run models at scale with strong price-performance and low operational friction
  • In this role, you will partner with engineering, product, infrastructure, and go-to-market teams to drive programs that improve how inference services are launched, onboarded, operated, and optimized
  • This includes managing the execution of customer-facing onboarding programs, launch readiness for dedicated inference capabilities, and the platform improvements needed to support scale, reliability, and predictable delivery
  • Drive end-to-end program management for inference platform initiatives spanning reliability, customer onboarding, launch readiness, and runtime optimization
  • Lead cross-functional programs for customer onboarding across dedicated and serverless inference offerings, ensuring clear ownership, launch criteria, and readiness for strategic customer use cases
  • Drive launch readiness for new inference capabilities by aligning teams around real customer outcomes, supportability, and end-to-end validation
  • Partner with engineering and product to define and deliver roadmap outcomes for latency, throughput, uptime, operational quality, and price-performance
  • Coordinate multi-team execution across platform, infrastructure, and customer-facing teams to deliver reliable and scalable inference services
  • Build and operationalize success metrics, dashboards, launch gates, and review cadences to measure service reliability, onboarding readiness, efficiency, and quality across the inference stack
  • Establish repeatable processes for release validation, performance regression tracking, launch management, and postmortem follow-through
  • Help unify operational processes, support mechanisms, and execution visibility across inference deployment models and customer onboarding paths
  • Create strong communication channels between Engineering, Product, Infrastructure, and Go-to-Market teams to align priorities and deliver predictable, high-impact outcomes
Requirements

Strong technical fluency in distributed inference systems, GPU compute, cloud-native architectures, and performance optimization

Wondering if you’re a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren’t a 100% skill or experience match. Here are a few qualities we’ve found compatible with our team. If some of this describes you, we’d love to talk

  • Understanding of customer onboarding for technical products, especially where platform capabilities, infrastructure readiness, and support processes must align for launch
  • Bachelor’s degree in a technical field or equivalent practical experience
  • Proven experience driving large-scale infrastructure or platform programs from concept to production in complex, cross-functional environments
  • Excellent written and verbal communication skills, with the ability to align engineering, product, infrastructure, and customer-facing stakeholders around shared goals
  • Experience with inference-serving systems, model onboarding workflows, rollout strategies, and observability tooling
  • Familiarity with launch readiness, supportability, incident follow-through, and release validation for production infrastructure or platform services
  • 8+ years of technical program management experience in distributed systems, cloud infrastructure, or AI/ML platform engineering
  • Demonstrated success driving measurable improvements in reliability, performance, operational readiness, or customer delivery
  • Experience operating in high-growth environments where roadmap execution, reliability expectations, and customer commitments must be managed in parallel
  • You’re effective at creating clarity and momentum across ambiguous, fast-moving multi-team initiatives
  • You love driving execution for complex, customer-facing AI infrastructure and platform programs
  • You’re curious about how large-scale inference systems evolve across runtime performance, operational excellence, and customer onboarding
  • You enjoy turning technically complex platform work into predictable execution and successful launches
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Program Manager
Technical Program Manager

CoreWeave • York and North Yorkshire

On-site
GBP 90,000 - 135,000
Technical Program Manager (Performance & Benchmarking)
Technical Program Manager (Performance & Benchmarking)

CoreWeave • York and North Yorkshire

On-site
GBP 90,000 - 130,000
Technical Program Manager (Platform)
Technical Program Manager (Platform)

Scale AI • York and North Yorkshire

Hybrid
GBP 90,000 - 130,000
Health insurance
Dental & Vision
Mental healthcare
+1
Lead Technical Program Manager, AI Platform
Lead Technical Program Manager, AI Platform

Wayve • Greater London

Hybrid
GBP 120,000 - 180,000
Hybrid work policy
Office in London
Lead, AI Inference Platform Programs
Lead, AI Inference Platform Programs

CoreWeave • York and North Yorkshire

On-site
GBP 80,000 - 110,000
Technical Program Manager, Platform
Technical Program Manager, Platform

Front Door Defense • Greater London

On-site
GBP 90,000 - 120,000
Senior Software Engineer (Inference)
Senior Software Engineer (Inference)

AssemblyAI • Greater London

Remote
GBP 90,000 - 150,000
Home office stipend
Equity grant
Premium medical, dental, vision plans
+4
Technical Program Manager, Platform
Technical Program Manager, Platform

Scale • Greater London

On-site
GBP 95,000 - 150,000
Technical Program Manager, Platform London, UK Apply →
Technical Program Manager, Platform London, UK Apply →

Scale AI, Inc. • Greater London

On-site
GBP 90,000 - 150,000
Senior Product Manager, Data Infrastructure
Senior Product Manager, Data Infrastructure

Jobtailor • Greater London

Hybrid
GBP 90,000 - 140,000