GPU Kernel Engineer — Fast ML Training

MakerMaker.AI

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

MakerMaker.AI in San Francisco is seeking a skilled Software Engineer to write and optimize GPU kernels. You will work on deep low-level tasks that directly impact the performance of machine learning models.

The ideal candidate has over 4 years of experience with GPU kernels, strong systems expertise, and a proven track record in kernel optimizations. This role requires on-site work in a collaborative environment.

Qualifications

  • 4+ years writing performant GPU kernels (CUDA, ROCm, Triton, or equivalent).
  • Fluency with memory hierarchy, occupancy, and tensor cores.
  • Previous experience shipping kernel-level optimizations.

Responsibilities

  • Write and optimize GPU kernels for training and inference workloads.
  • Profile workloads and translate findings into kernel optimizations.
  • Integrate optimized kernels into training and serving stacks.

Skills

Performance optimization
CUDA
Profiling tools
Systems expertise
Python
C++

Tools

Nsight
ncu

Job description

MakerMaker.AI in San Francisco is seeking a skilled Software Engineer to write and optimize GPU kernels. You will work on deep low-level tasks that directly impact the performance of machine learning models.

The ideal candidate has over 4 years of experience with GPU kernels, strong systems expertise, and a proven track record in kernel optimizations. This role requires on-site work in a collaborative environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Kernel Engineer — Fast ML Training & Inference
GPU Kernel Engineer — Fast ML Training & Inference

Tilde Research • Palo Alto (CA)

On-site
USD 150,000 - 260,000
GPU Kernel Engineer — High-Performance ML at Scale
GPU Kernel Engineer — High-Performance ML at Scale

The Consensus • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation with equity
100% medical, dental, and vision insurance coverage
Flexible PTO policy including a Winter Break
+2
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
GPU Kernel Engineer: Build Fast AI Inference at Scale
GPU Kernel Engineer: Build Fast AI Inference at Scale

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
Founding GPU Kernel Engineer — Hand‑Tuned ML Kernels
Founding GPU Kernel Engineer — Hand‑Tuned ML Kernels

San Francisco Tensor Company • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Equity
Comprehensive benefits package
CUDA GPU Kernel Architect for High-Throughput AI
CUDA GPU Kernel Architect for High-Throughput AI

TypeSafe AI • San Francisco (CA)

On-site
USD 180,000 - 280,000
Base salary $180k–$280k plus equity
100% covered health insurance
Daily lunch and dinner
+2
KERNEL ENGINEER
KERNEL ENGINEER

MakerMaker.AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Pioneering GPU Kernel Engineer for ML Performance
Pioneering GPU Kernel Engineer for ML Performance

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000
GPU Kernel Engineer: Performance-Driven ML
GPU Kernel Engineer: Performance-Driven ML

Mindbeam • United States

On-site
USD 100,000 - 140,000
GPU Kernel Maestro for AI Training & Inference
GPU Kernel Maestro for AI Training & Inference

Advanced Micro Devices, Inc. • Bellevue (WA)

On-site
USD 180,000 - 240,000
AMD benefits