Applied AI Inference Performance Engineer

TheDataJob

San Francisco (CA)

On-site

USD 188,000 - 275,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision insurance
ESPP – Employee Stock Purchase Program
Tuition Reimbursement
Paid Parental Leave
Flexible PTO
Casual work environment

Job summary

CoreWeave, The Essential Cloud for AI, seeks an Applied AI Engineer to measure and improve the real-world performance of its inference platform. You will build benchmarks, profile model-serving behavior, and drive optimizations across platforms, runtimes, and hardware toward scalable, cost-efficient AI workloads.

The role emphasizes applied performance work with collaboration across teams. Expect to contribute through experiments, validation, and reproducible workflows.

Qualifications

  • 4+ years of experience in machine learning, systems, performance engineering, or adjacent applied engineering work.

Responsibilities

  • Build and maintain benchmarking workflows measuring latency, throughput, and cost.
  • Benchmark inference stack against real customer workloads and baselines.
  • Profile model-serving behavior across frameworks, runtimes, and hardware to find bottlenecks.
  • Drive targeted optimization for customer workloads and validate changes with traces and benchmarks.
  • Design experiments on model-serving techniques like quantization and caching strategies.

Skills

Python programming
Performance benchmarking
Empirical evaluation
Model serving experience
LLM inference systems
Cross-functional collaboration

Tools

Nsight Systems
PyTorch profilers
TensorRT-LLM
vLLM
SGLang

Job description

CoreWeave, The Essential Cloud for AI, seeks an Applied AI Engineer to measure and improve the real-world performance of its inference platform. You will build benchmarks, profile model-serving behavior, and drive optimizations across platforms, runtimes, and hardware toward scalable, cost-efficient AI workloads.

The role emphasizes applied performance work with collaboration across teams. Expect to contribute through experiments, validation, and reproducible workflows.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Applied AI Inference Performance Engineer
Applied AI Inference Performance Engineer

CoreWeave • San Francisco (CA)

On-site
USD 188,000 - 275,000
Medical benefits
401(k) with match
Flexible PTO
+5
Inference Performance Engineer - Benchmark & Optimize
Inference Performance Engineer - Benchmark & Optimize

CoreWeave • Sunnyvale (CA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
401(k) with employer match
Paid parental leave
+2
Applied AI Inference Engineer - Benchmark & Optimize
Applied AI Inference Engineer - Benchmark & Optimize

CoreWeave • Bellevue (WA)

On-site
USD 188,000 - 275,000
Medical Insurance
Dental Insurance
Vision Insurance
+12
Inference Performance Engineer: Benchmark & Optimize
Inference Performance Engineer: Benchmark & Optimize

Coreweave • Bellevue (CA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
Equity awards
401(k) with match
+1
Senior AI Infra Performance & Observability Engineer
Senior AI Infra Performance & Observability Engineer

CoreWeave • Sunnyvale (CA)

On-site
USD 182,000 - 242,000
Medical, dental, vision insurance
401(k) with employer match
Paid Parental Leave
+1
Senior Performance Engineer for AI Platform
Senior Performance Engineer for AI Platform

Weights & Biases • Bellevue (WA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+3
Senior AI Infrastructure Performance & Observability Engineer
Senior AI Infrastructure Performance & Observability Engineer

CoreWeave • Bellevue (WA)

On-site
USD 182,000 - 242,000
Health insurance
Life Insurance
Disability insurance
+5
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
Staff Performance Engineer: Scale AI Platforms
Staff Performance Engineer: Scale AI Platforms

Weights & Biases • San Francisco (CA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
Equity awards
401(k) with match
+2
Senior GPU Kernel Engineer for High-Performance AI Inference
Senior GPU Kernel Engineer for High-Performance AI Inference

CoreWeave • Sunnyvale (CA)

On-site
USD 182,000 - 242,000
Medical, dental, and vision insurance
401(k) with employer match
ESPP
+2