Technical Program Manager, Inference Platform Partnerships

Apply

San Francisco, Northern (CA, KY)

Hybrid

USD 130,000 - 170,000

Full time

11 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Perplexity is seeking a technical program manager to bridge model providers, engineering, and product teams, driving the core inference platform forward from roadmap to production. You will orchestrate across external partners and internal groups to keep new models and capacity moving smoothly, while building a scalable operating model for this growing function.

The role requires strong technical judgment, cross-team coordination, and a hands-on approach to balancing latency, throughput, uptime,

Qualifications

  • Strong experience with technical program management or product management in infrastructure, distributed systems, or ML/model-serving products.
  • Direct experience with production LLM or ML inference, understanding latency, reliability and cost.
  • Comfort orchestrating across external partners and internal engineering teams with competing priorities and timelines.
  • Experience with data and metrics, and the judgment to surface difficult tradeoffs between latency, throughput, uptime, and cost.
  • Thrives in a small, agile team; has initiative and desire for ownership without much precedent.
  • 6+ years of combined technical program management or product management experience.

Responsibilities

  • Execute the roadmap for the inference platform - request handling, rate limits and quotas, usage controls, and the reliability and observability surface engineering and product teams depend on
  • Be the connective tissue between model providers and Perplexity's engineering and product teams - coordinating onboarding, launch readiness, and rollout for new models and capacity
  • Drive latency, throughput, uptime, and cost-efficiency as core execution metrics, surfacing tradeoffs between them rather than letting them become side effects
  • Run the operating model for model-release and optimization programs, including day-zero launches, across performance engineering, infrastructure, and product teams
  • Lead cross-functional delivery for inference-stack changes, from planning through launch and post-launch validation
  • Build the mechanisms that make releases predictable - rituals, dashboards, launch checklists - so inference releases stay low-risk at Perplexity's scale
  • Partner with GPU capacity and compute teams to reconcile execution decisions against cost, capacity, and vendor constraints

Skills

Technical program management
ML inference
Cross-team coordination
Data & metrics
Ownership mindset
6+ years experience

Job description

Perplexity is seeking a technical program manager to bridge model providers, engineering, and product teams, driving the core inference platform forward from roadmap to production. You will orchestrate across external partners and internal groups to keep new models and capacity moving smoothly, while building a scalable operating model for this growing function.

The role requires strong technical judgment, cross-team coordination, and a hands-on approach to balancing latency, throughput, uptime,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Program Manager, Inference Platform
Technical Program Manager, Inference Platform

Perplexity • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Technical Program Manager - AI Inference Platform
Technical Program Manager - AI Inference Platform

Applied Methods Ltd • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Member of Technical Staff (TPM, Inference)
Member of Technical Staff (TPM, Inference)

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 130,000 - 170,000
Member of Technical Staff (TPM, Inference)
Member of Technical Staff (TPM, Inference)

Perplexity • San Francisco (CA), Northern (KY)

On-site
USD 180,000 - 240,000
1d Perplexity Member of Technical Staff (TPM, Inference) San Francisco Perplexity 1d Member of Technical Staff (TPM, Inference) San Francisco
1d Perplexity Member of Technical Staff (TPM, Inference) San Francisco Perplexity 1d Member of Technical Staff (TPM, Inference) San Francisco

Applied Methods Ltd • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Staff Software Engineer - Model Serving & Systems
Staff Software Engineer - Model Serving & Systems

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Technical Program Manager — AI Inference Platform
Technical Program Manager — AI Inference Platform

Baseten • United States

Remote
USD 140,000 - 170,000
Staff Engineer, AI Model-Serving Platform
Staff Engineer, AI Model-Serving Platform

Perplexity • San Francisco (CA)

On-site
USD 220,000 - 405,000
Inference Platform TPM — Lead Scalable AI Delivery
Inference Platform TPM — Lead Scalable AI Delivery

CoreWeave • Bellevue (WA)

On-site
USD 198,000 - 264,000
Medical benefits
401(k) with match
ESPP
+4
Member of Technical Staff (Software Engineer, Model Platform)
Member of Technical Staff (Software Engineer, Model Platform)

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 260,000