Staff Engineer, AI Inference Optimization

DigitalOcean

Boston (MA)

On-site

USD 191,200 - 239,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity compensation
Bonus potential
Conference reimbursement
LinkedIn Learning access
Flexible time off

Job summary

DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. Responsible for architectural decisions maximizing throughput and minimizing latency for large models, you will lead performance strategies across GPU clusters and guide the technical roadmap.

You will mentor through code reviews, collaborate with product teams, and push cutting-edge optimization techniques to keep DigitalOcean at the forefront of Gen AI performance.

Qualifications

  • 5+ years in high-performance computing or AI infrastructure.
  • Deep familiarity with Gen AI landscape and model families.
  • Hands-on optimization of attention layers and distributed GPUs.
  • Understanding of NVIDIA/AMD GPUs and software stacks.
  • Experience integrating and contributing to open-source projects.
  • Strong system design for low-level GPU programming and memory access.
  • Technical lead experience guiding design and delivery.

Responsibilities

  • Lead benchmarking and performance optimization at engine and kernel layers.
  • Engineer solutions for attention, memory, and FP8/BF16 tradeoffs.
  • Identify kernel fusion opportunities and optimize across multi-node GPUs.
  • Advise on hardware procurement and software integration.
  • Mentor through code reviews and drive cross-functional alignment.
  • Collaborate with product managers to translate hardware limits into features.

Skills

Technical Depth
Gen AI Literacy
Optimization Expert
Hardware Fluency
Open Source Mastery
Systems Design
Leadership through Influence
Low-Level Mastery
CUDA/Triton

Tools

CUDA
ROCm
TensorRT
OpenAI Triton
Triton compiler

Job description

DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. Responsible for architectural decisions maximizing throughput and minimizing latency for large models, you will lead performance strategies across GPU clusters and guide the technical roadmap.

You will mentor through code reviews, collaborate with product teams, and push cutting-edge optimization techniques to keep DigitalOcean at the forefront of Gen AI performance.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Inference Optimization Engineer
Senior AI Inference Optimization Engineer

DigitalOcean • Seattle (WA)

On-site
USD 191,000 - 239,000
Senior AI Inference Optimization Engineer
Senior AI Inference Optimization Engineer

DigitalOcean • Austin (TX)

On-site
USD 191,200 - 239,000
Senior AI Inference Optimization Engineer (Remote)
Senior AI Inference Optimization Engineer (Remote)

DigitalOcean • United States

Remote
USD 191,000 - 239,000
Remote Senior AI Inference Optimization Engineer
Remote Senior AI Inference Optimization Engineer

DigitalOcean • San Francisco (CA)

On-site
USD 191,200 - 239,000
Equity compensation
Remote work
Senior AI Inference Performance Engineer (Remote)
Senior AI Inference Performance Engineer (Remote)

DigitalOcean • Denver (CO)

On-site
USD 191,200 - 239,000
Staff Engineer, Inference Optimizations
Staff Engineer, Inference Optimizations

DigitalOcean • United States

Remote
USD 191,000 - 239,000
Staff Engineer, Inference Optimizations
Staff Engineer, Inference Optimizations

DigitalOcean • Seattle (WA)

On-site
USD 191,000 - 239,000
Senior Engineer 2: GPU Kernel and Performance
Senior Engineer 2: GPU Kernel and Performance

DigitalOcean • San Francisco (CA)

On-site
USD 167,200 - 209,000
Competitive salary
Flexible time off policy
Employee Assistance Program
+2
Senior Engineer II — Serverless AI Inference (Hybrid)
Senior Engineer II — Serverless AI Inference (Hybrid)

DigitalOcean • Seattle (WA)

Hybrid
USD 167,000 - 209,000
Equity compensation
Bonus potential
Senior Engineering Manager, AI Inference & Kubernetes
Senior Engineering Manager, AI Inference & Kubernetes

DigitalOcean, LLC • Seattle (WA)

Hybrid
USD 200,800 - 251,000
Equity compensation
Education reimbursement
Flexible time-off policy