Senior AI Inference & Optimization Infrastructure Engineer

DiDi

San Jose (CA)

On-site

USD 170,000 - 351,000

Full time

45 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DiDi Autonomous Driving seeks a Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization, to lead performance tuning and deployment of AI models across vehicle-side and cloud infrastructure.

You will design high-efficiency inference pipelines, build system-level stability frameworks, and optimize hardware execution for ultra-low latency in autonomous systems. You will collaborate with cross-functional teams to accelerate model iteration and ensure rock-solid, production-grade

Qualifications

  • Master's degree or higher in a closely related technical field.
  • 3+ years in high-performance computing, AI infrastructure, or embedded deployment.
  • Strong skills in C++ and Python with CUDA/OpenMP experience.
  • Familiar with TensorRT/ONNX Runtime and LLM inference frameworks.
  • Knowledge of modern GPU architectures and memory bandwidth management.

Responsibilities

  • Own deployment, optimization, and resource scheduling of vehicle-side AI models.
  • Lead stability initiatives and root-cause analysis for performance bottlenecks.
  • Architect and scale deployment environments for LLMs and foundational models.
  • Track industry methodologies and implement advanced optimization toolchains.
  • Establish profiling/telemetry using CUDA tools across GPU architectures.
  • Collaborate with perception/prediction, cloud infra, and safety teams.

Skills

C++
Python
CUDA
OpenMP
Profiling tools

Education

Master's in Computer Science

Tools

TensorRT
ONNX Runtime
vLLM
SGLang
TensorRT-LLM

Job description

DiDi Autonomous Driving seeks a Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization, to lead performance tuning and deployment of AI models across vehicle-side and cloud infrastructure.

You will design high-efficiency inference pipelines, build system-level stability frameworks, and optimize hardware execution for ultra-low latency in autonomous systems. You will collaborate with cross-functional teams to accelerate model iteration and ensure rock-solid, production-grade

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Infrastructure Engineer, Inference & Optimization
Senior AI Infrastructure Engineer, Inference & Optimization

Didi Labs • San Jose (CA)

On-site
USD 170,000 - 351,000
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization

Didi Labs • San Jose (CA)

On-site
USD 170,000 - 351,000
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization

DiDi • San Jose (CA)

On-site
USD 170,000 - 351,000
Senior AI Infra Engineer: High-Performance Inference
Senior AI Infra Engineer: High-Performance Inference

Ddn • Sacramento (CA)

On-site
USD 140,000 - 200,000
Staff Engineer, AI Inference Optimization
Staff Engineer, AI Inference Optimization

DigitalOcean • Boston (MA)

On-site
USD 191,000 - 239,000
Equity compensation
Bonus potential
Conference reimbursement
+2
Remote Senior AI Inference Optimization Engineer
Remote Senior AI Inference Optimization Engineer

DigitalOcean • San Francisco (CA)

On-site
USD 191,000 - 239,000
Equity compensation
Remote work
Senior AI Inference Data Plane Engineer
Senior AI Inference Data Plane Engineer

DigitalOcean • Seattle (WA)

Hybrid
USD 139,000 - 174,000
Equity compensation
Hybrid work model
Senior Systems Engineer, AI Inference Infra
Senior Systems Engineer, AI Inference Infra

United States Digital Space LLC • United States

Remote
USD 150,000 - 190,000
Senior AI Inference Optimizations Engineer — Remote
Senior AI Inference Optimizations Engineer — Remote

DigitalOcean • Seattle (WA)

Remote
USD 167,000 - 209,000
Senior AI Inference Optimization Engineer
Senior AI Inference Optimization Engineer

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 124,000 - 196,000
Equity
Benefits package
Hybrid work model