Senior Director, AI Inference & Optimization

DigitalOcean

San Francisco (CA)

On-site

USD 274,400 - 343,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DigitalOcean is seeking a Senior Director of Engineering to lead a high‑performing team building and scaling our LLM inference products across control plane, optimization, and architecture. You will own Serverless Inference, Dedicated Inference, Inference Router, Batch Inference, and Multimodal Inference, driving cost‑efficient systems for millions of users worldwide.

You will lead product strategy, collaborate with PMs and engineers, and champion observability, performance, and reliability

Qualifications

  • 10+ years of software engineering experience.
  • 6+ years in a technical leadership or management role.
  • Experience in Inference Systems or AI/ML systems preferred.

Responsibilities

  • Lead and mentor an engineering team focused on LLM inference products.
  • Define and execute the product roadmap for DigitalOcean's inference suite.
  • Drive the design of the inference serving stack, optimization, and model architecture.
  • Collaborate with Product, other engineering teams, and stakeholders on priorities and risks.
  • Ensure production health, SLAs, and on-call processes.

Skills

Team Leadership
Distributed systems
Kubernetes at scale
LLM inference
AI workload orchestration
Communication
Ownership

Tools

vLLM
SGLang
LLM-D
NVLink

Job description

DigitalOcean is seeking a Senior Director of Engineering to lead a high‑performing team building and scaling our LLM inference products across control plane, optimization, and architecture. You will own Serverless Inference, Dedicated Inference, Inference Router, Batch Inference, and Multimodal Inference, driving cost‑efficient systems for millions of users worldwide.

You will lead product strategy, collaborate with PMs and engineers, and champion observability, performance, and reliability

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Head of AI Inference Platforms & Optimizations
Head of AI Inference Platforms & Optimizations

DigitalOcean • Seattle (WA)

Hybrid
USD 274,000 - 343,000
Senior AI Inference Data Plane Engineer
Senior AI Inference Data Plane Engineer

DigitalOcean • Seattle (WA)

Hybrid
USD 139,000 - 174,000
Equity compensation
Hybrid work model
Staff Engineer, AI Inference Optimization
Staff Engineer, AI Inference Optimization

DigitalOcean • Boston (MA)

On-site
USD 191,000 - 239,000
Equity compensation
Bonus potential
Conference reimbursement
+2
Senior Engineering Manager, AI Inference & Kubernetes
Senior Engineering Manager, AI Inference & Kubernetes

DigitalOcean, LLC • Seattle (WA)

Hybrid
USD 200,000 - 251,000
Equity compensation
Education reimbursement
Flexible time-off policy
Senior Director, Inference Products and Optimizations
Senior Director, Inference Products and Optimizations

DigitalOcean • Seattle (WA)

Hybrid
USD 274,000 - 343,000
Senior Engineer 2: AI Inference Engine Systems
Senior Engineer 2: AI Inference Engine Systems

DigitalOcean • San Francisco (CA)

On-site
USD 167,000 - 209,000
Competitive salary
Career development opportunities
Comprehensive benefits package
+2
Senior Engineer 2: AI Inference Engine Systems
Senior Engineer 2: AI Inference Engine Systems

DigitalOcean • Seattle (WA)

On-site
USD 167,000 - 209,000
Career development resources
Competitive benefits package
Equity compensation options
Senior AI Inference Optimization Engineer
Senior AI Inference Optimization Engineer

DigitalOcean • Austin (TX)

On-site
USD 191,000 - 239,000
Senior AI Inference Performance Engineer (Remote)
Senior AI Inference Performance Engineer (Remote)

DigitalOcean • Denver (CO)

On-site
USD 191,000 - 239,000
Lead, AI Inference & GPU Strategy
Lead, AI Inference & GPU Strategy

DigitalOcean • Seattle (WA)

Hybrid
USD 218,000 - 273,000