Senior AI Benchmarking & Performance Engineer

CoreWeave

Sunnyvale (CA)

On-site

USD 182,000 - 242,000

Full time

24 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Medical, dental, and vision insurance
401(k) with generous employer match
Tuition Reimbursement
Flexible PTO
Catered lunch every day

Job summary

CoreWeave is seeking a Senior Engineer for the Benchmarking & Performance team to own the design and delivery of planet-scale performance data workflows. You will drive latency, throughput, and reliability improvements across multiple services while partnering with product, orchestration, and hardware teams to meet strict P99 SLAs at scale.

You will develop benchmarks and workflows for MLPerf training and inference, contribute to architecture decisions, and mentor junior engineers.

Qualifications

  • 3–5 years of distributed systems experience.
  • Strong Python or Go programming.
  • Production Kubernetes experience.
  • Familiarity with CI/CD and observability tools.
  • Exposure to GPU systems or model-serving stacks.

Responsibilities

  • Develop and enhance Kubernetes-native benchmarking services measuring latency, throughput, jitter, and cost-per-request.
  • Implement benchmarking workflows for MLPerf Training and Inference runs, including workload setup and validation.
  • Participate in design discussions and architecture decisions.
  • Break down tasks into milestones and deliver high-quality code.
  • Maintain reproducible benchmarking processes and documentation.
  • Provide code reviews and share best practices with peers.
  • Mentor junior engineers and elevate coding/testing standards.

Skills

Python
Go
Distributed systems
Kubernetes
CI/CD
Performance tuning

Tools

Prometheus
Grafana
OpenTelemetry
CUDA
NCCL
Megatron-LM

Job description

CoreWeave is seeking a Senior Engineer for the Benchmarking & Performance team to own the design and delivery of planet-scale performance data workflows. You will drive latency, throughput, and reliability improvements across multiple services while partnering with product, orchestration, and hardware teams to meet strict P99 SLAs at scale.

You will develop benchmarks and workflows for MLPerf training and inference, contribute to architecture decisions, and mentor junior engineers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Benchmarking & Performance Architect
AI Benchmarking & Performance Architect

CoreWeave • Sunnyvale (CA)

On-site
USD 206,000 - 333,000
Medical, dental, and vision insurance
Equity awards
401(k) with match
+2
Applied AI Inference Engineer: Benchmark & Optimize
Applied AI Inference Engineer: Benchmark & Optimize

CoreWeave • San Francisco (CA)

On-site
USD 188,000 - 275,000
Medical insurance
Life Insurance
Disability insurance
+5
Inference Performance Engineer — Applied AI
Inference Performance Engineer — Applied AI

CoreWeave • Seattle (WA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
Equity awards
Discretionary bonus
+6
AI/ML Platform TPM: Performance & Benchmarking Lead
AI/ML Platform TPM: Performance & Benchmarking Lead

CoreWeave • San Francisco (CA)

On-site
USD 177,000 - 237,000
Medical, dental, vision insurance
401(k) with employer match
Flexible PTO
+5
Senior CUDA Kernel Engineer for High-Performance Inference
Senior CUDA Kernel Engineer for High-Performance Inference

CoreWeave • Bellevue (WA)

On-site
USD 182,000 - 242,000
Medical, dental, and vision insurance
Company-paid Life Insurance
Disability insurance
+3
Senior Software Engineer - Perf and Benchmarking
Senior Software Engineer - Perf and Benchmarking

CoreWeave • Sunnyvale (CA)

On-site
USD 182,000 - 242,000
Medical, dental, and vision insurance
401(k) with generous employer match
Tuition Reimbursement
+2
AI/ML Platform TPM: Performance & Benchmarking Lead
AI/ML Platform TPM: Performance & Benchmarking Lead

CoreWeave • Sunnyvale (CA)

On-site
USD 177,000 - 237,000
Medical, dental, and vision insurance
401(k) with employer match
Flexible PTO
+3
AI/ML Platform TPM — Performance & Benchmark Lead
AI/ML Platform TPM — Performance & Benchmark Lead

CoreWeave • Livingston (NJ)

On-site
USD 177,000 - 237,000
Medical, dental, and vision insurance
Equity awards
401(k) match
+6
Senior Performance Engineer - AI Fabric & GPU Scale
Senior Performance Engineer - AI Fabric & GPU Scale

Astera Labs • San Jose (CA)

On-site
USD 135,000 - 170,000
Senior Data Platform Engineer — AI-Scale Pipelines
Senior Data Platform Engineer — AI-Scale Pipelines

Weights & Biases • New York (NY)

On-site
USD 165,000 - 242,000
Medical, dental, vision insurance
401(k) with employer match
Flexible PTO
+2