Technical Program Manager, Inference Platform

Perplexity

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 240,000

Full time

11 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Perplexity is seeking a technical program manager to unite model providers, engineering, and product teams, driving our inference platform forward. You will orchestrate across partners and internal teams to keep new models and capacity moving smoothly into production, while executing the roadmap for the inference platform itself.

The ideal candidate has strong technical judgment, thrives coordinating across external partners with competing timelines, and is energized by building scalable

Qualifications

  • 6+ years of combined technical program management or product management experience.
  • Direct experience with production LLM or ML inference.
  • Comfort orchestrating across external partners and internal engineering teams.
  • Experience with data and metrics, and the judgment to surface tradeoffs between latency, throughput, uptime, and cost.
  • Thrives in a small, agile team; has initiative and desire for ownership without much precedent to lean on.

Responsibilities

  • Execute the roadmap for the inference platform.
  • Coordinate onboarding, launch readiness, and rollout for new models and capacity.
  • Drive latency, throughput, uptime, and cost-efficiency.
  • Run the operating model for model-release and optimization programs.
  • Lead cross-functional delivery for inference-stack changes.
  • Build mechanisms that make releases predictable with rituals, dashboards, and launch checklists.
  • Partner with GPU capacity and compute teams to reconcile execution decisions against cost and constraints.

Skills

Technical PM
Product management
Distributed systems
ML model-serving
Cross-functional coordination
Metrics/data analysis

Job description

Perplexity is seeking a technical program manager to unite model providers, engineering, and product teams, driving our inference platform forward. You will orchestrate across partners and internal teams to keep new models and capacity moving smoothly into production, while executing the roadmap for the inference platform itself.

The ideal candidate has strong technical judgment, thrives coordinating across external partners with competing timelines, and is energized by building scalable

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Program Manager, Inference Platform Partnerships
Technical Program Manager, Inference Platform Partnerships

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 130,000 - 170,000
Technical Program Manager - AI Inference Platform
Technical Program Manager - AI Inference Platform

Applied Methods Ltd • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Member of Technical Staff (TPM, Inference)
Member of Technical Staff (TPM, Inference)

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 130,000 - 170,000
Member of Technical Staff (TPM, Inference)
Member of Technical Staff (TPM, Inference)

Perplexity • San Francisco (CA), Northern (KY)

On-site
USD 180,000 - 240,000
1d Perplexity Member of Technical Staff (TPM, Inference) San Francisco Perplexity 1d Member of Technical Staff (TPM, Inference) San Francisco
1d Perplexity Member of Technical Staff (TPM, Inference) San Francisco Perplexity 1d Member of Technical Staff (TPM, Inference) San Francisco

Applied Methods Ltd • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Staff Software Engineer - Model Serving & Systems
Staff Software Engineer - Model Serving & Systems

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Technical Program Manager — AI Inference Platform
Technical Program Manager — AI Inference Platform

Baseten • United States

Remote
USD 140,000 - 170,000
Staff Engineer, AI Model-Serving Platform
Staff Engineer, AI Model-Serving Platform

Perplexity • San Francisco (CA)

On-site
USD 220,000 - 405,000
Member of Technical Staff (Software Engineer, Model Platform)
Member of Technical Staff (Software Engineer, Model Platform)

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Inference Platform TPM — Lead Scalable AI Delivery
Inference Platform TPM — Lead Scalable AI Delivery

CoreWeave • Bellevue (WA)

On-site
USD 198,000 - 264,000
Medical benefits
401(k) with match
ESPP
+4