GPU Kernel Engineer for High-Performance AI Inference

Baseten

San Francisco (CA)

On-site

USD 180,000 - 360,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
Paid parental leave
Fertility and family-building stipend
Company-facilitated 401(k)

Job summary

Baseten in San Francisco is seeking a GPU Kernel Engineer to enhance AI acceleration by designing high-performance GPU kernels. This role involves optimizing ML operations, utilizing CUDA and various performance profiling tools. Candidates should have a strong understanding of GPU architecture and proficient coding skills in C++. The position offers competitive compensation, including equity, with a full benefits package including 100% medical coverage for employees and their dependents.

Qualifications

  • Strong understanding of GPU architecture and programming paradigms.
  • Proficient in C++ and GPU performance profiling tools.
  • Knowledge of numerical precision and quantization strategies.

Responsibilities

  • Design and implement high-performance GPU kernels for key ML operations.
  • Write and optimize code using CUDA and architecture‑specific techniques.
  • Identify and resolve performance bottlenecks using profiling tools.

Skills

Strong understanding of GPU architecture
Proficient in C++
Experience with CUDA C++ API
Knowledge of memory access patterns

Tools

CUDA
Nsight Systems
Torch Profiler

Job description

Baseten in San Francisco is seeking a GPU Kernel Engineer to enhance AI acceleration by designing high-performance GPU kernels. This role involves optimizing ML operations, utilizing CUDA and various performance profiling tools. Candidates should have a strong understanding of GPU architecture and proficient coding skills in C++. The position offers competitive compensation, including equity, with a full benefits package including 100% medical coverage for employees and their dependents.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Kernel Engineer for AI Inference & Performance
GPU Kernel Engineer for AI Inference & Performance

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Flexible working hours
Daily lunch and dinner
Health check-up support
+3
GPU Kernel Engineer: Build Fast AI Inference at Scale
GPU Kernel Engineer: Build Fast AI Inference at Scale

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • New York (NY)

On-site
USD 180,000 - 360,000
Senior AI Kernel Engineer — GPU Inference & Kernel Optimization
Senior AI Kernel Engineer — GPU Inference & Kernel Optimization

Modular • United States

Hybrid
USD 198,000 - 286,000
Senior GPU Inference Engine Engineer
Senior GPU Inference Engine Engineer

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Flexible working hours
Daily lunch and dinner provided; unlimited snacks and beverages
Health check-up support and top-tier equipment/hardware support
+2
Staff GenAI Kernel & Performance Engineer
Staff GenAI Kernel & Performance Engineer

Databricks • San Francisco (CA)

On-site
USD 190,000 - 233,000
CUDA GPU Kernel Architect for High-Throughput AI
CUDA GPU Kernel Architect for High-Throughput AI

TypeSafe AI • San Francisco (CA)

On-site
USD 180,000 - 280,000
Base salary $180k–$280k plus equity
100% covered health insurance
Daily lunch and dinner
+2
Pioneering GPU Kernel Engineer for ML Performance
Pioneering GPU Kernel Engineer for ML Performance

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000
Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Kernel & Compiler Performance Engineer (GPU/AI)
Senior Kernel & Compiler Performance Engineer (GPU/AI)

RadixArk • Palo Alto (CA)

On-site
Competitive compensation
Comprehensive benefits
Flexible work arrangements