Principal Product Manager, Inference Engine

DigitalOcean

Seattle (WA)

Hybrid

USD 218,400 - 273,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DigitalOcean is seeking a Senior Product Manager to own the GPU strategy for AI inference, aligning product and infrastructure to optimize GPU capacity, pricing, and deployment. You will partner with engineering to shape the roadmap for caching, autoscaling, batching, and latency, while balancing developer needs with cost and performance.

You will collaborate with AI-native startups, mid-market, and strategic accounts to translate customer needs into scalable product initiatives and optimized

Qualifications

  • Experience building infrastructure or AI platforms for developers and enterprises.
  • Strong understanding of GPU economics, capacity planning and cost-to-serve.
  • Familiarity with modern AI workloads including LLM inference and model serving.

Responsibilities

  • Own the GPU strategy and pricing across serverless and dedicated inference offerings.
  • Define the product roadmap with engineering for prompt caching, autoscaling, batching, latency, and observability.
  • Balance developer experience with infrastructure economics and cost-to-serve.
  • Create pricing models and packaging for serverless, dedicated, and enterprise inference.
  • Drive customer-informed product decisions with AI-native startups and strategic accounts.
  • Collaborate with engineering and infra teams on GPU fleet planning and capacity allocation.
  • Establish operating metrics: GPU utilization, latency, revenue per GPU hour, margin.

Skills

GPU economics
LLM inference
Product strategy
Executive communication
Analytical rigor
Ownership mindset
Customer obsession
Infrastructure experience
Cross-functional collaboration

Job description

Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized by the fast-paced environment of a true industry disruptor, you’ll find your place here. We value winning together—while learning, having fun, and making a profound difference for the dreamers and builders in the world.

Are you passionate about building the infrastructure that will power the next generation of AI applications? Are you ready to own the GPU strategy behind one of DigitalOcean’s fastest-growing product categories?

DigitalOcean is entering a pivotal moment as we build infrastructure for AI-native companies and the future 100 million developers. Inference is becoming one of the most important layers of the AI stack. Developers need access to the right models, on the right GPUs, with the right latency, pricing, reliability, and scale. At the same time, cloud providers must manage scarce GPU capacity with discipline: maximizing utilization, improving gross margin, selecting the right model mix, and ensuring customers can trust the platform for production workloads.

What You’ll Do
  • Own the GPU strategy for the inference business: Define how DigitalOcean should deploy, allocate, price, and optimize GPU capacity across serverless inference, dedicated inference, batch workloads, and future inference offerings.
  • Maximize GPU utilization and margin: Create a clear product and business framework for improving token revenue per GPU hour, reducing idle capacity, reclaiming underutilized infrastructure, and driving better gross margin as the business scales.
  • Define the inference product roadmap: Partner with engineering to prioritize capabilities such as prompt caching, autoscaling, batching, latency optimization, observability, dedicated deployments, compliance features, and media model support.
  • Balance developer experience with infrastructure economics: Build products that are simple for developers to use while making rigorous tradeoffs around latency, availability, throughput, pricing, and cost-to-serve.
  • Create pricing and packaging strategy: Work with finance, GTM, and engineering to define SKUs, pricing models, discounting frameworks, and packaging for serverless, dedicated, and enterprise inference customers.
  • Drive customer‑backed product decisions: Work directly with AI‑native startups, mid‑market customers, and strategic accounts to understand model needs, performance requirements, compliance expectations, and deployment patterns.
  • Partner deeply with engineering and infrastructure teams: Translate customer demand and business goals into infrastructure requirements across GPU fleet planning, model serving, capacity allocation, performance optimization, and reliability.
  • Establish operating metrics for the business: Define and track the metrics that matter, including GPU utilization, token throughput, revenue per GPU hour, latency, error rates, model adoption, margin, customer retention, and capacity efficiency.
What You’ll Bring
  • Deep product judgment in infrastructure or AI: Experience building infrastructure, developer platforms, ML platforms, inference systems, cloud services, or highly technical products for developers and enterprises.
  • Strong understanding of GPU economics: Ability to reason about utilization, throughput, latency, CapEx, cost‑to‑serve, gross margin, capacity planning, and workload placement.
  • Fluency in modern AI workloads: Familiarity with LLM inference, open‑source models, model serving, prompt caching, batching, model routing, media models, latency tradeoffs, and production AI application patterns.
  • Technical depth with business orientation: You can work credibly with infrastructure engineers while also making clear product and business tradeoffs for executives, GTM teams, and customers.
  • Strong analytical rigor: You are comfortable building frameworks, models, and decision systems that turn ambiguous infrastructure and customer signals into clear product direction.
  • Customer obsession: You work backwards from developers and AI‑native companies, but you also understand that great infrastructure products must be reliable, performant, simple, and economically sustainable.
  • Executive communication: You can explain complex technical and business decisions clearly to senior leaders, customers, and cross‑functional teams.
  • Ownership mindset: You thrive in ambiguous, fast‑moving environments where the product category is still forming and the right answer requires judgment, experimentation, and operational discipline.
Compensation Range
  • $218,400 - $273,000
  • Hybrid role

DigitalOcean is an equal‑opportunity employer. We do not discriminate on the basis of race, religion, color, ancestry, national origin, caste, sex, sexual orientation, gender, gender identity or expression, age, disability, medical condition, pregnancy, genetic makeup, marital status, or military service.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Director, Inference Products and Optimizations
Senior Director, Inference Products and Optimizations

DigitalOcean • Seattle (WA)

Hybrid
USD 274,000 - 343,000
Senior Engineering Manager, Kernel and Virt
Senior Engineering Manager, Kernel and Virt

DigitalOcean, LLC • Seattle (WA)

Hybrid
USD 200,000 - 251,000
Equity compensation
Education reimbursement
Flexible time-off policy
Senior Engineer 2: GPU Kernel and Performance
Senior Engineer 2: GPU Kernel and Performance

DigitalOcean • Seattle (WA)

On-site
USD 167,000 - 209,000
Flexible time off
Employee Stock Purchase Program
Career development resources
Senior Engineer 2: AI Inference Engine Systems
Senior Engineer 2: AI Inference Engine Systems

DigitalOcean • San Francisco (CA)

On-site
USD 167,000 - 209,000
Competitive salary
Career development opportunities
Comprehensive benefits package
+2
Senior Engineer 2: AI Inference Engine Systems
Senior Engineer 2: AI Inference Engine Systems

DigitalOcean • Seattle (WA)

On-site
USD 167,000 - 209,000
Career development resources
Competitive benefits package
Equity compensation options
Staff Engineer, Inference Optimizations
Staff Engineer, Inference Optimizations

DigitalOcean • Seattle (WA)

Hybrid
USD 191,000 - 239,000
Senior Engineer, Inference Data Plane
Senior Engineer, Inference Data Plane

DigitalOcean • Seattle (WA)

Hybrid
USD 139,000 - 174,000
Equity compensation
Hybrid work model
Senior Engineer 2: GPU Kernel and Performance
Senior Engineer 2: GPU Kernel and Performance

DigitalOcean • San Francisco (CA)

On-site
USD 167,000 - 209,000
Competitive salary
Flexible time off policy
Employee Assistance Program
+2
Staff Forward Deployed Engineer
Staff Forward Deployed Engineer

3M HEALTHCARE • Seattle (WA)

On-site
USD 195,000 - 239,000
Competitive salary
Flexible time off policy
Equity compensation
+1
Staff Product Manager
Staff Product Manager

DigitalOcean • Seattle (WA)

Hybrid
USD 186,000 - 233,000
Career development reimbursement
Flexible time off policy
Equity compensation options