Technical Program Manager, Model Performance

Jobtailor

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 240,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Jobtailor in San Francisco is looking for a Technical Program Manager to own execution across the Model Performance project portfolio, designing planning structures, operating cadences, and status reporting mechanisms.

You will coordinate end-to-end model release and optimization programs, surface risks early, and drive cross-team alignment as the scope expands from Model Performance Core into Model APIs and the inference production stack, while partnering with engineering leads to scale the

Qualifications

  • Running programs of this scope at architecture-evel complexity
  • Experience program-managing model performance or inference optimization
  • Understanding vLLM, TensorRT-LLM, SGLang, NVIDIA Dynamo in production
  • Comfort with ambiguity and zero-to-one program building
  • Ability to influence without authority across engineers, managers, and leadership
  • Excellent written and verbal communication
  • High-agency decision making with ownership and accountability

Responsibilities

  • Own execution across Model Performance's active project portfolio
  • Design and establish planning structures, operating cadences, and status reporting mechanisms
  • Coordinate model release and optimization programs end to end, including day-zero launches
  • Sequence work across performance engineering, infrastructure, and release stakeholders
  • Drive cross-team alignment as scope expands from Model Performance Core into Model APIs and the inference production stack (BIS)
  • Surface risks and dependencies early and keep leadership informed with clear, honest status
  • Partner with engineering leads to design team structures and ownership boundaries as the organization scales

Skills

Technical PM
Model Performance Optimization
Cross-Team Alignment
Influencing Without Authority
High-Agency Decision Making
Excellent Communication
Ownership & Accountability

Tools

vLLM
TensorRT-LLM
SGLang
NVIDIA Dynamo

Job description

  • • Own execution across Model Performance's active project portfolio
  • • Design and establish planning structures, operating cadences, and status reporting mechanisms
  • • Coordinate model release and optimization programs end to end, including day-zero launches
  • • Sequence work across performance engineering, infrastructure, and release stakeholders
  • • Drive cross-team alignment as scope expands from Model Performance Core into Model APIs and the inference production stack (BIS)
  • • Surface risks and dependencies early and keep leadership informed with clear, honest status
  • • Partner with engineering leads to design team structures and ownership boundaries as the organization scales
Requirements
  • Deep technical program management experience; already running programs of this scope at an organization of similar or greater complexity
  • Experience program-managing model performance or inference optimization work
  • Understanding of how vLLM, TensorRT-LLM, SGLang, or NVIDIA Dynamo fit into a production serving stack
  • Comfort with ambiguity and zero-to-one program building
  • Ability to influence without authority across engineers, managers, and leadership
  • Excellent written and verbal communication
  • High-agency decision making demonstrating ownership, accountability, and a strong desire to get things done
Core Competencies

Demonstrates deep technical program management expertise with a focus on model performance and inference optimization. Capable of driving cross-team alignment and managing complex projects while maintaining clear communication and accountability.

Highest-signal resume keywords
  • Technical Program Management
  • Model Performance Optimization
  • Cross-Team Alignment
  • Influencing Without Authority
  • High-Agency Decision Making
ATS Optimization Keywords
Hard Skills
  • Program Management
  • Model Optimization
  • Project Execution
  • Risk Management
  • Status Reporting
Soft Skills
  • Excellent Communication
  • Comfort with Ambiguity
  • Ownership
  • Accountability
Industry Keywords
  • Model Performance
  • Inference Production Stack
  • Performance Engineering
  • Project Portfolio Management
Tools & Technologies
  • VLLM
  • TensorRT-LLM
  • SGLang
  • NVIDIA Dynamo
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Program Manager, Model Performance
Technical Program Manager, Model Performance

Baseten • New York (NY)

On-site
USD 165,000 - 330,000
Equity
Medical, dental, vision insurance
Flexible PTO
+3
Technical Program Manager, Model Performance
Technical Program Manager, Model Performance

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Equity
Medical coverage
Flexible PTO
+3
Technical Program Manager- LLM supporting Nvidia
Technical Program Manager- LLM supporting Nvidia

Sustainable Talent • California (MO)

On-site
USD 272,214,000 - 358,176,000
Technical Program Manager, Model Performance
Technical Program Manager, Model Performance

Baseten • San Francisco (CA), Northern (KY)

On-site
USD 170,000 - 230,000
Competitive compensation
Equity
Medical/dental/vision insurance
+4
Technical Program Manager- LLM supporting Nvidia
Technical Program Manager- LLM supporting Nvidia

Sustainable Talent • Santa Clara (CA)

On-site
USD 131,000 - 172,000
Full benefits
PTO
Member of Technical Staff – Software Engineer
Member of Technical Staff – Software Engineer

Jobtailor • Palo Alto (CA)

On-site
USD 180,000 - 280,000
Machine Learning Engineer, LLM Inference Optimization
Machine Learning Engineer, LLM Inference Optimization

GMI Cloud • San Francisco (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff, ML Performance
Member of Technical Staff, ML Performance

Odyssey • Palo Alto (CA)

On-site
USD 130,000 - 160,000
Technical Program Manager – AI/ML
Technical Program Manager – AI/ML

DynPro Inc. • Santa Clara (CA)

On-site
USD 140,000 - 210,000
Senior Technical Program Manager (Engineering) - AI Tooling & Systems
Senior Technical Program Manager (Engineering) - AI Tooling & Systems

Madrona Venture Labs • United States

On-site
USD 150,000 - 230,000