AI Performance & Kernel Engineer for Frontier-Scale ML

Zyphra

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Comprehensive medical, dental, vision, and FSA plans
Competitive compensation and 401(k) plan
Relocation and immigration support
In-office snacks and meals
Unlimited PTO and company holidays
Collaborative high-energy environment

Job summary

A leading AI technology firm located in San Francisco is seeking a Research Engineer specializing in AI Performance & Kernel Optimization. The role involves enhancing the performance of large-scale AI systems, optimizing kernels, and collaborating with various teams. Ideal candidates should have a strong engineering background in GPU kernel development and experience with ML workloads. Benefits include comprehensive health plans, competitive compensation, and an engaging work environment.

Qualifications

  • Experience writing highly performant GPU kernels.
  • Experience optimizing ML workloads for large-scale training.
  • Strong understanding of distributed training systems.

Responsibilities

  • Improve performance of language model training and inference stacks.
  • Optimize kernel development for ML workloads.
  • Collaborate with research and infrastructure teams.

Skills

Building reliable, high-performance systems
Low-level performance intuition
Excellent communication skills
Collaboration across teams
Performance tuning expertise

Education

Background in physics, mathematics, computer science, or electrical engineering

Tools

PTX
CUDA
HIP
Triton

Job description

A leading AI technology firm located in San Francisco is seeking a Research Engineer specializing in AI Performance & Kernel Optimization. The role involves enhancing the performance of large-scale AI systems, optimizing kernels, and collaborating with various teams. Ideal candidates should have a strong engineering background in GPU kernel development and experience with ML workloads. Benefits include comprehensive health plans, competitive compensation, and an engaging work environment.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Kernel Performance Engineer - AI Tooling & Systems
Kernel Performance Engineer - AI Tooling & Systems

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Kernel Performance Engineer for AI Tooling
Kernel Performance Engineer for AI Tooling

OpenAI • Los Angeles (CA)

On-site
USD 100,000 - 130,000
Staff GenAI Kernel & Performance Engineer
Staff GenAI Kernel & Performance Engineer

Databricks • San Francisco (CA)

On-site
USD 190,900 - 232,800
Annual performance bonus
Equity options
Comprehensive benefits package
GPU Kernel Engineer for AI Inference & Performance
GPU Kernel Engineer for AI Inference & Performance

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Flexible working hours
Daily lunch and dinner
Health check-up support
+3
GPU Kernel Engineer for High-Performance AI Inference
GPU Kernel Engineer for High-Performance AI Inference

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
GPU Kernel Engineer: Build Fast AI Inference at Scale
GPU Kernel Engineer: Build Fast AI Inference at Scale

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
Senior Kernel & Compiler Performance Engineer (GPU/AI)
Senior Kernel & Compiler Performance Engineer (GPU/AI)

RadixArk • Palo Alto (CA)

On-site
USD 210,000 - 290,000
Competitive compensation
Comprehensive benefits
Flexible work arrangements
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads

VeeAR Projects Inc. • Sunnyvale (CA)

On-site
USD 140,000 - 210,000
Software Engineer - ML Model Performance
Software Engineer - ML Model Performance

Baseten • San Francisco (CA)

On-site
USD 150,000 - 250,000
Competitive compensation with equity
100% medical, dental, and vision insurance
Generous PTO policy
+2