Senior AI Systems Engineer — High-Performance Inference

DDN

United States

Remote

USD 150,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DDN seeks an engineer to build and optimize production AI systems, focusing on fast inference across GPU and CPU paths. You’ll tackle bottlenecks in memory, storage, and KV caching while designing scalable infrastructure for RAG and retrieval‑heavy workloads.

You will own critical areas from architecture decisions to hands‑on implementation, operating in technically demanding environments where efficiency and scale determine success. PhD preferred but practical impact matters most.

Qualifications

  • Experience building or optimizing production AI systems, not only experimentation.
  • Understand how inference performance is shaped by compute, memory, storage, and serving architecture.
  • Hands‑on work close to the systems layer, optimizing workloads across GPU/CPU and reducing bottlenecks.
  • Real ownership in areas like model serving, retrieval, caching, storage, or distributed performance.
  • Ability to move between architecture decisions and hands‑on implementation in scalable environments.
  • Background in AI infrastructure, high‑performance systems, or adjacent distributed systems.

Responsibilities

  • Build and optimize LLM serving and inference systems for production environments.
  • Improve performance across GPU and CPU pathways.
  • Work on KV cache, memory, storage, and throughput bottlenecks.
  • Design and scale systems that support RAG and retrieval‑heavy AI workloads.
  • Contribute to infrastructure where storage architecture and systems efficiency affect AI performance.
  • Solve engineering problems at the intersection of AI, high‑performance systems, and distributed infrastructure.

Skills

Production AI systems experience
Inference performance optimization
Systems-level engineering
Ownership in model serving/retrieval/c
Architecture-to-implementation fluency
AI infrastructure or HPC background

Education

PhD preferred

Tools

GPU/CPU resource tuning
Distributed storage systems

Job description

DDN seeks an engineer to build and optimize production AI systems, focusing on fast inference across GPU and CPU paths. You’ll tackle bottlenecks in memory, storage, and KV caching while designing scalable infrastructure for RAG and retrieval‑heavy workloads.

You will own critical areas from architecture decisions to hands‑on implementation, operating in technically demanding environments where efficiency and scale determine success. PhD preferred but practical impact matters most.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Infra Engineer: High-Performance Inference
Senior AI Infra Engineer: High-Performance Inference

Ddn • Sacramento (CA)

On-site
USD 140,000 - 200,000
Senior AI Infrastructure Engineer – High-Performance Serving
Senior AI Infrastructure Engineer – High-Performance Serving

Data Direct Networks • California (MO)

On-site
USD 150,000 - 230,000
Vacation plans
Paid holidays
Bonus programs
+5
Senior AI Data Path & Storage Architect
Senior AI Data Path & Storage Architect

DDN • California (MO)

Hybrid
USD 180,000 - 250,000
Senior AI Infrastructure Engineer — Production Systems
Senior AI Infrastructure Engineer — Production Systems

DDN • San Francisco (CA)

On-site
USD 150,000 - 250,000
Senior AI Data Path & Storage Architect
Senior AI Data Path & Storage Architect

Data Direct Networks • California (MO)

On-site
USD 180,000 - 240,000
Highly Competitive Vacation Plans
Paid Holidays
Bonus Programs
+5
Senior/Staff AI Engineer
Senior/Staff AI Engineer

DDN • San Francisco (CA)

On-site
USD 150,000 - 250,000
Senior Staff Engineer - AI Data Path
Senior Staff Engineer - AI Data Path

DDN • California (MO)

Hybrid
USD 180,000 - 250,000
Senior Staff Engineer - AI Data Path
Senior Staff Engineer - AI Data Path

Data Direct Networks • California (MO)

On-site
USD 180,000 - 240,000
Highly Competitive Vacation Plans
Paid Holidays
Bonus Programs
+5
Senior AI Systems Engineer: GPU Kernels & Inference Equity
Senior AI Systems Engineer: GPU Kernels & Inference Equity

NVIDIA AI • Michigan

On-site
USD 150,000 - 190,000
Equity
Health Insurance
Remote Senior AI Inference Optimization Engineer
Remote Senior AI Inference Optimization Engineer

DigitalOcean • San Francisco (CA)

On-site
USD 191,000 - 239,000
Equity compensation
Remote work