GPU Kernel Engineer

TypeSafe AI

San Francisco (CA)

On-site

USD 180,000 - 280,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Base salary $180k–$280k plus equity
100% covered health insurance
Daily lunch and dinner
Visa sponsorships
401K plans

Job summary

TypeSafe AI in San Francisco is seeking a GPU kernel engineer with deep CUDA expertise to accelerate our training and inference workloads. You will write, optimize, and maintain high-performance kernels close to the metal.

Work alongside research and platform engineers to squeeze throughput and reduce latency, with the team operating in-person at our SF office near Embarcadero. This role requires hands-on LLM training experience and a track record of real performance wins.

Qualifications

  • Deep CUDA / GPU kernel expertise.
  • Experience building and optimizing inference / training kernels.
  • Hands-on LLM training experience (not hobbyist).
  • Reason from first principles about performance, memory, and parallelism.

Responsibilities

  • Write, optimize, and maintain high-performance GPU kernels (e.g., in CUDA / CuTe DSL) for training and inference
  • Profile end-to-end performance and eliminate bottlenecks across the stack
  • Partner with research and platform engineers to squeeze maximum throughput and minimum latency out of our hardware

Skills

CUDA / GPU kernel expertise
Inference & training kernels
LLM training experience
Performance optimization

Tools

CuTe DSL
CUDA

Job description

About TypeSafe

TypeSafe is a frontier model lab. We build reliable and general AI systems to power economically valuable automation. Our mission is to usher in a new era of Transformative Artificial Intelligence (TAI): technology with the power to drive a societal shift on the scale of the agricultural and industrial revolutions.

While others chase benchmarks and academic puzzles, we've been quietly rethinking the LLM stack from first principles — building a new kind of general frontier model designed for real-world reliability, decision-making, and autonomy in production.

We're a small, fast-moving team from OpenAI, Google Brain, and Meta/FAIR, backed by top-tier investors. Since mid-2024, we've been engineering the foundation for what comes after the current "state-of-the-art" — a model that actually gets things done.

About the role

We're looking for a GPU kernel engineer with deep, low-level CUDA expertise to make our training and inference faster and more efficient. You'll write and optimize custom kernels, profile and eliminate bottlenecks, and work close to the metal across our model stack.

Responsibilities include:

  • Write, optimize, and maintain high-performance GPU kernels (e.g., in CUDA / CuTe DSL) for training and inference

  • Profile end-to-end performance and eliminate bottlenecks across the stack

  • Partner with research and platform engineers to squeeze maximum throughput and minimum latency out of our hardware

We are looking for people who
  • Have deep CUDA / GPU kernel expertise and a track record of real performance wins

  • Have built and optimized inference / training kernels

  • Have hands-on LLM training experience (real, not at a hobbyist level)

  • Reason from first principles about performance, memory, and parallelism

  • Are responsible, ownership-inclined team players who are mission aligned and excited to go all-in

Life at TypeSafe

We're a small, flat, close-knit team dedicated to real-world impact preparing the world for Transformative AI. Our team works fully in-person in our San Francisco office near Embarcadero station. We love what we do and care about our work a lot.

We strive for excellence and craftsmanship and won't stop until we get there. When the team wins, we all win, and we enjoy collaborating and inspiring each other to grow as a team and as individuals.

We also value emotional honesty, kindness, and bringing your whole self to work. We build machines; we don't try to be machines.

We want you to be able to do the most impactful work of your career at TypeSafe and help define our future as a company.

We provide
  • Base salary of $180k–280k plus equity, based on leveling

  • 100% covered health insurance

  • Daily lunch and dinner

  • Visa sponsorships

  • 401K plans

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Infrastructure Engineer, Kubernetes Specialist
Infrastructure Engineer, Kubernetes Specialist

TypeSafe AI • San Francisco (CA)

On-site
USD 180,000 - 280,000
Base salary of $180k-280k plus equity
100% covered health insurance
Daily lunch and dinner
+2
Founding GPU Kernel Engineer
Founding GPU Kernel Engineer

San Francisco Tensor Company • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Equity
Comprehensive benefits package
Engineering Baseten San Francisco
Engineering Baseten San Francisco

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 230,000 - 320,000
Competitive equity
100% medical, dental, vision
Flexible PTO including Winter Break
+4
Product Engineer
Product Engineer

TypeSafe AI • San Francisco (CA)

On-site
USD 180,000 - 280,000
180k-280k base salary plus equity
100% covered health insurance
Daily lunch and dinner
+2
Software Engineer - GPU Kernels
Software Engineer - GPU Kernels

The Consensus • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation with equity
100% medical, dental, and vision insurance coverage
Flexible PTO policy including a Winter Break
+2
Software Engineer - GPU Kernels
Software Engineer - GPU Kernels

Baseten • New York (NY)

On-site
USD 180,000 - 360,000
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
Paid parental leave
+2
Software Engineer - GPU Kernels
Software Engineer - GPU Kernels

Baseten • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
100% medical coverage
Generous PTO policy
+2
Founding GPU Kernel Engineer
Founding GPU Kernel Engineer

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000
Software Engineer - GPU Kernels
Software Engineer - GPU Kernels

Baseten • San Francisco (CA)

On-site
USD 180,000 - 360,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
+3
Research Engineer, Infrastructure, Kernels
Research Engineer, Infrastructure, Kernels

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Generous health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1