Staff GenAI Kernel & Performance Engineer

Databricks

San Francisco (CA)

On-site

USD 190,900 - 232,800

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Annual performance bonus
Equity options
Comprehensive benefits package

Job summary

A leading data and AI company in San Francisco seeks a Staff Software Engineer to lead kernel-level performance engineering for GenAI workloads. The role involves designing and optimizing high-performance GPU kernels, mentoring engineers, and driving performance roadmaps for low-level compute paths. Ideal candidates should have advanced experience with GPU architectures and performance optimization techniques. This position offers a competitive salary and a chance to work with a talented team focused on pushing the frontier of inference performance.

Qualifications

  • Experience tuning compute kernels for ML workloads.
  • Strong knowledge of GPU architecture and memory hierarchy.
  • Familiarity with ML-specific kernel libraries.

Responsibilities

  • Lead the design and implementation of compute kernels.
  • Drive performance improvements and kernel optimizations.
  • Collaborate with teams to roll out optimizations in production.

Skills

CUDA
GPU acceleration
Deep learning
Performance optimization
Debugging

Education

BS/MS/PhD in Computer Science

Tools

Nsight
NVProf

Job description

A leading data and AI company in San Francisco seeks a Staff Software Engineer to lead kernel-level performance engineering for GenAI workloads. The role involves designing and optimizing high-performance GPU kernels, mentoring engineers, and driving performance roadmaps for low-level compute paths. Ideal candidates should have advanced experience with GPU architectures and performance optimization techniques. This position offers a competitive salary and a chance to work with a talented team focused on pushing the frontier of inference performance.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
GPU Kernel Engineer for AI Inference & Performance
GPU Kernel Engineer for AI Inference & Performance

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Flexible working hours
Daily lunch and dinner
Health check-up support
+3
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
GPU Kernel Engineer: Build Fast AI Inference at Scale
GPU Kernel Engineer: Build Fast AI Inference at Scale

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
Senior AI Kernel Engineer — GPU Inference & Kernel Optimization
Senior AI Kernel Engineer — GPU Inference & Kernel Optimization

Modular • United States

Hybrid
USD 198,000 - 286,000
GenAI Inference Optimization Lead — GPU Performance
GenAI Inference Optimization Lead — GPU Performance

Advanced Micro Devices • San Jose (CA)

Hybrid
USD 150,000 - 200,000
AI Performance & Kernel Engineer for Frontier-Scale ML
AI Performance & Kernel Engineer for Frontier-Scale ML

Zyphra • San Francisco (CA)

On-site
USD 120,000 - 160,000
Comprehensive medical, dental, vision, and FSA plans
Competitive compensation and 401(k) plan
Relocation and immigration support
+3
Senior Kernel & Compiler Performance Engineer (GPU/AI)
Senior Kernel & Compiler Performance Engineer (GPU/AI)

RadixArk • Palo Alto (CA)

On-site
USD 210,000 - 290,000
Competitive compensation
Comprehensive benefits
Flexible work arrangements
Kernel Performance Engineer - AI Tooling & Systems
Kernel Performance Engineer - AI Tooling & Systems

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior GPU Inference Engine Engineer
Senior GPU Inference Engine Engineer

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Flexible working hours
Daily lunch and dinner provided; unlimited snacks and beverages
Health check-up support and top-tier equipment/hardware support
+2