Staff Engineer, AI Inference Optimization

DigitalOcean

Boston (MA)

On-site

USD 191,200 - 239,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity compensation
Bonus potential
Conference reimbursement
LinkedIn Learning access
Flexible time off

Job summary

DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. Responsible for architectural decisions maximizing throughput and minimizing latency for large models, you will lead performance strategies across GPU clusters and guide the technical roadmap.

You will mentor through code reviews, collaborate with product teams, and push cutting-edge optimization techniques to keep DigitalOcean at the forefront of Gen AI performance.

Qualifications

  • 5+ years in high-performance computing or AI infrastructure.
  • Deep familiarity with Gen AI landscape and model families.
  • Hands-on optimization of attention layers and distributed GPUs.
  • Understanding of NVIDIA/AMD GPUs and software stacks.
  • Experience integrating and contributing to open-source projects.
  • Strong system design for low-level GPU programming and memory access.
  • Technical lead experience guiding design and delivery.

Responsibilities

  • Lead benchmarking and performance optimization at engine and kernel layers.
  • Engineer solutions for attention, memory, and FP8/BF16 tradeoffs.
  • Identify kernel fusion opportunities and optimize across multi-node GPUs.
  • Advise on hardware procurement and software integration.
  • Mentor through code reviews and drive cross-functional alignment.
  • Collaborate with product managers to translate hardware limits into features.

Skills

Technical Depth
Gen AI Literacy
Optimization Expert
Hardware Fluency
Open Source Mastery
Systems Design
Leadership through Influence
Low-Level Mastery
CUDA/Triton

Tools

CUDA
ROCm
TensorRT
OpenAI Triton
Triton compiler

Job description

DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. Responsible for architectural decisions maximizing throughput and minimizing latency for large models, you will lead performance strategies across GPU clusters and guide the technical roadmap.

You will mentor through code reviews, collaborate with product teams, and push cutting-edge optimization techniques to keep DigitalOcean at the forefront of Gen AI performance.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Inference Optimization Engineer
Senior AI Inference Optimization Engineer

DigitalOcean • Austin (TX)

On-site
USD 191,000 - 239,000
Remote Senior AI Inference Optimization Engineer
Remote Senior AI Inference Optimization Engineer

DigitalOcean • San Francisco (CA)

On-site
USD 191,000 - 239,000
Equity compensation
Remote work
Senior AI Inference Performance Engineer (Remote)
Senior AI Inference Performance Engineer (Remote)

DigitalOcean • Denver (CO)

On-site
USD 191,000 - 239,000
Senior Engineer 2: GPU Kernel and Performance
Senior Engineer 2: GPU Kernel and Performance

DigitalOcean • Seattle (WA)

On-site
USD 167,000 - 209,000
Flexible time off
Employee Stock Purchase Program
Career development resources
Senior AI Inference Data Plane Engineer
Senior AI Inference Data Plane Engineer

DigitalOcean • Seattle (WA)

Hybrid
USD 139,000 - 174,000
Equity compensation
Hybrid work model
Staff Engineer, Inference Optimizations
Staff Engineer, Inference Optimizations

DigitalOcean • Seattle (WA)

Hybrid
USD 191,000 - 239,000
Senior Engineer 2: GPU Kernel and Performance
Senior Engineer 2: GPU Kernel and Performance

DigitalOcean • San Francisco (CA)

On-site
USD 167,000 - 209,000
Competitive salary
Flexible time off policy
Employee Assistance Program
+2
Senior Director, AI Inference & Optimization
Senior Director, AI Inference & Optimization

DigitalOcean • San Francisco (CA)

On-site
USD 274,000 - 343,000
Head of AI Inference Platforms & Optimizations
Head of AI Inference Platforms & Optimizations

DigitalOcean • Seattle (WA)

Hybrid
USD 274,000 - 343,000
Senior Engineer 2: AI Inference Engine Systems
Senior Engineer 2: AI Inference Engine Systems

DigitalOcean • Seattle (WA)

On-site
USD 167,000 - 209,000
Career development resources
Competitive benefits package
Equity compensation options