Senior GPU Kernel Engineer, High-Performance Compiler

San Francisco Tensor Company

San Francisco (CA)

On-site

USD 285,000 - 315,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Relocation assistance
Equity

Job summary

SF Tensor in San Francisco is hiring a Member of Technical Staff for GPU Kernel Engineering to push the limits of what the hardware can do before any search begins.

You will hand-write and optimize kernels to achieve unprecedented throughput, profile at the microarchitectural level, reason about PTX and SASS, and build models that feed the compiler search space. Relocation assistance is offered and equity accompanies competitive compensation.

Qualifications

  • Hand-write kernels that beat vendor libraries.
  • Reading PTX, SASS, or equivalent machine-level ISA.
  • Low-level profiling to optimize kernel throughput.

Responsibilities

  • Write and hand-optimize kernels for real workloads.
  • Profile microarchitectural performance and identify bottlenecks.
  • Develop machine-searchable structures for compiler exploration.
  • Work below PTX at the ISA level to optimize schedules.

Skills

Kernel optimization
C++
CUDA
Profiling
GPU architectures

Tools

Nsight Compute
Nsight Systems
rocprof
omniperf

Job description

SF Tensor in San Francisco is hiring a Member of Technical Staff for GPU Kernel Engineering to push the limits of what the hardware can do before any search begins.

You will hand-write and optimize kernels to achieve unprecedented throughput, profile at the microarchitectural level, reason about PTX and SASS, and build models that feed the compiler search space. Relocation assistance is offered and equity accompanies competitive compensation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer — AI-Driven GPU Compiler Optimization
Staff Engineer — AI-Driven GPU Compiler Optimization

SF Tensor • San Francisco (CA)

On-site
USD 275,000 - 315,000
Relocation assistance
Equity and benefits
Office in San Francisco
Senior GPU Compiler Engineer — End-to-End MLIR & CUDA
Senior GPU Compiler Engineer — End-to-End MLIR & CUDA

San Francisco Tensor Company • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Office in San Francisco
Member of Technical Staff, GPU Kernels
Member of Technical Staff, GPU Kernels

San Francisco Tensor Company • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Equity
Pioneering GPU Kernel Engineer for ML Performance
Pioneering GPU Kernel Engineer for ML Performance

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000
Senior GPU Compiler Engineer (MLIR/LLVM)
Senior GPU Compiler Engineer (MLIR/LLVM)

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Member of Technical Staff, GPU Compiler
Member of Technical Staff, GPU Compiler

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Member of Technical Staff, GPU Compiler
Member of Technical Staff, GPU Compiler

San Francisco Tensor Company • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Office in San Francisco
Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior GPU Compiler & Performance Engineer
Senior GPU Compiler & Performance Engineer

AMD • San Jose (CA)

On-site
USD 180,000 - 300,000
Founding GPU Kernel Engineer
Founding GPU Kernel Engineer

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000