Senior AI Systems Performance Engineer

Doist

San Jose (CA)

On-site

USD 140,000 - 190,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
HSA contributions
Dental
Vision
Disability insurance
Life insurance
FSA
Headspace

Job summary

SambaNova Systems in San Jose is seeking a talented ML performance engineer to optimize and scale state-of-the-art foundation models on our reconfigurable dataflow platform. You will work hands-on with leading models to push throughput, latency, and efficiency, bridging gaps between compiler, runtime, and hardware teams to deliver world-record AI inference performance.

The role requires a strong foundation in deep learning and systems optimization, 3+ years of experience, and proficiency in

Qualifications

  • Bachelor's or higher in computer science, electrical engineering, or a related field (e.g., applied mathematics, physics, or statistics).
  • 3+ years of experience in Deep learning model development and performance optimization, compiler/runtime/kernel-level optimization, or software–hardware co-design.
  • Proficiency in Python or C++, with strong foundations in algorithms, data structures, and numerical computing.
  • Experience with at least one major ML framework — PyTorch, TensorFlow, or JAX.
  • Demonstrated ability to analyze and optimize performance in real-world ML pipelines.

Responsibilities

  • Bring up and optimize cutting-edge foundation models (e.g., DeepSeek, Llama, Qwen, and others) on the SambaNova platform through the SambaNova software stack.
  • Profile and enhance model performance across compiler, runtime, and hardware layers to achieve SOTA throughput and latency.
  • Collaborate with machine learning, compiler, runtime, and hardware teams to deliver co-designed, high-performance AI applications.
  • Integrate the latest advances in model architecture, quantization, scheduling, and memory optimization from both academia and industry.
  • Develop robust, scalable, and efficient end-to-end inference solutions aligned with customer needs.
  • Identify performance bottlenecks and propose dataflow or scheduling optimizations for both single-node and distributed systems.

Skills

Python
C++
PyTorch
TensorFlow
JAX

Education

Bachelor's degree in CS/EE or related

Tools

CUDA
Triton
DeepSpeed
Megatron
vLLM
TensorRT
CuDNN

Job description

SambaNova Systems in San Jose is seeking a talented ML performance engineer to optimize and scale state-of-the-art foundation models on our reconfigurable dataflow platform. You will work hands-on with leading models to push throughput, latency, and efficiency, bridging gaps between compiler, runtime, and hardware teams to deliver world-record AI inference performance.

The role requires a strong foundation in deep learning and systems optimization, 3+ years of experience, and proficiency in

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Inference Performance Engineer
Senior AI Inference Performance Engineer

SambaNova • San Jose (CA)

On-site
USD 180,000 - 240,000
AI Systems Performance Engineer - Scalable LLM Inference
AI Systems Performance Engineer - Scalable LLM Inference

SambaNova Systems • San Jose (CA)

On-site
USD 135,000 - 165,000
Equity
Health insurance
Well-being benefits
Senior AI Systems Performance Engineer San Jose, California, United States
Senior AI Systems Performance Engineer San Jose, California, United States

SambaNova • Palo Alto (CA)

On-site
USD 120,000 - 150,000
95% premium coverage for employee medical insurance
Health Savings Account with employer contribution
Flexible Spending Account options
Principal AI Systems Performance Engineer
Principal AI Systems Performance Engineer

SambaNova • San Jose (CA)

On-site
USD 180,000 - 240,000
Principal AI Systems Performance Engineer
Principal AI Systems Performance Engineer

Socket.dev • San Jose (CA)

On-site
USD 140,000 - 190,000
Health insurance
HSA contributions
Dental
+5
Senior ML Infra Engineer: High-Throughput AI Inference
Senior ML Infra Engineer: High-Throughput AI Inference

SambaNovaSystems • United States

On-site
USD 200,000 - 275,000
Health insurance
Health Savings Account (HSA)
Headspace subscription
+2
Senior AI Systems Performance Engineer: Drive SOTA Inference
Senior AI Systems Performance Engineer: Drive SOTA Inference

SambaNova • Palo Alto (CA)

On-site
USD 120,000 - 150,000
95% premium coverage for employee medical insurance
Health Savings Account with employer contribution
Flexible Spending Account options
Senior AI Inference Platform Engineer
Senior AI Inference Platform Engineer

SambaNova • San Jose (CA)

On-site
USD 144,000 - 189,000
Health insurance
HSA
Gympass+
+1
Senior ML Performance Engineer: Scale & Throughput
Senior ML Performance Engineer: Scale & Throughput

NLP PEOPLE • Sunnyvale (CA)

On-site
USD 215,000 - 285,000
AI Systems Performance Engineer - New Graduate
AI Systems Performance Engineer - New Graduate

SambaNova Systems • San Jose (CA)

On-site
USD 135,000 - 165,000
Equity
Health insurance
Well-being benefits