Senior AI Systems Performance Engineer: Drive SOTA Inference

SambaNova

Palo Alto (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

95% premium coverage for employee medical insurance
Health Savings Account with employer contribution
Flexible Spending Account options

Job summary

A leader in AI technology in Palo Alto is seeking a Senior AI Systems Performance Engineer to optimize the latest foundation models on their innovative platform. This role involves collaborating with cross-functional teams to push the performance limits of AI systems. Candidates should have a degree in computer science or related fields and experience in deep learning, performance optimization, and major ML frameworks. This position offers competitive salary and comprehensive benefits.

Qualifications

  • 3+ years of experience in deep learning model development.
  • Demonstrated ability to analyze and optimize performance in ML pipelines.
  • Experience with at least one major ML framework.

Responsibilities

  • Optimize and scale foundation models on SambaNova's platform.
  • Profile and enhance performance across compiler, runtime, and hardware layers.
  • Collaborate with teams to deliver high-performance AI applications.

Skills

Deep learning model development and performance optimization
Compiler, runtime, or kernel-level optimization
Proficiency in Python or C++
Analysis and optimization of performance in real-world ML pipelines

Education

Bachelor's or higher degree in computer science, electrical engineering, or a related field

Tools

PyTorch
TensorFlow
CUDA

Job description

A leader in AI technology in Palo Alto is seeking a Senior AI Systems Performance Engineer to optimize the latest foundation models on their innovative platform. This role involves collaborating with cross-functional teams to push the performance limits of AI systems. Candidates should have a degree in computer science or related fields and experience in deep learning, performance optimization, and major ML frameworks. This position offers competitive salary and comprehensive benefits.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Inference Performance Architect
Senior AI Inference Performance Architect

NVIDIA • California (MO)

On-site
USD 152,000 - 242,000
Senior AI Systems Performance Engineer
Senior AI Systems Performance Engineer

NVIDIA • Town of Texas (WI)

On-site
USD 184,000 - 288,000
Senior AI Systems Performance Engineer
Senior AI Systems Performance Engineer

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Senior AI Inference Performance Engineer (Remote)
Senior AI Inference Performance Engineer (Remote)

DigitalOcean • San Francisco (CA)

Remote
USD 167,000 - 209,000
Senior AI Inference Performance Architect | Equity Options
Senior AI Inference Performance Architect | Equity Options

NVIDIA Corporation • California (MO)

Hybrid
USD 152,000 - 241,500
Senior AI Model Serving Engineer — Low-Latency Inference
Senior AI Model Serving Engineer — Low-Latency Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 166,000 - 225,000
Annual performance bonus
Equity options
Comprehensive benefits package
AI Systems Performance Engineer
AI Systems Performance Engineer

OpenAI • Los Angeles (CA)

Hybrid
USD 100,000 - 150,000
Relocation assistance
Hybrid work model
Senior AI Systems Performance Engineer
Senior AI Systems Performance Engineer

OpenAI • United States

On-site
USD 325,000 - 405,000
GPU Performance Engineer: Scale AI Inference
GPU Performance Engineer: Scale AI Inference

Anthropic • San Francisco (CA)

On-site
USD 315,000 - 560,000
Competitive salary
Equity opportunities
Flexible working hours
+1
ML Performance Engineer – Real-Time Inference
ML Performance Engineer – Real-Time Inference

Odyssey • Palo Alto (CA)

On-site
USD 130,000 - 160,000