Founding ML Inference Performance Engineer

uRun

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health, dental, and vision
401(k) participation
Flexible spending accounts
Paid time off
Access to AI tools
MacBook Pro and AirPods

Job summary

uRun, located in San Francisco, is seeking a founding ML Performance Engineer to drive AI infrastructure performance. In this role, you will write custom CUDA kernels and optimize model inference for real-time applications, significantly impacting performance across the stack.

The ideal candidate will possess deep knowledge of CUDA, experience with AI workloads, and a strong capacity for optimization. The position offers a competitive salary, equity, and top-tier tools for an exceptional contributor.

Qualifications

  • Deep, hands-on CUDA expertise with production kernels.
  • Strong background in model inference at scale.
  • Fluency in hardware-aware algorithm design.

Responsibilities

  • Write custom CUDA kernels for performance improvements.
  • Optimize model inference targeting sub-50ms latency.
  • Profile and benchmark full inference pipeline.

Skills

CUDA expertise
Model inference optimization
GPU memory hierarchy knowledge
Benchmarking and profiling

Tools

PyTorch
TensorRT

Job description

uRun, located in San Francisco, is seeking a founding ML Performance Engineer to drive AI infrastructure performance. In this role, you will write custom CUDA kernels and optimize model inference for real-time applications, significantly impacting performance across the stack.

The ideal candidate will possess deep knowledge of CUDA, experience with AI workloads, and a strong capacity for optimization. The position offers a competitive salary, equity, and top-tier tools for an exceptional contributor.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Founding Real-Time AI Platform Engineer
Founding Real-Time AI Platform Engineer

uRun • San Francisco (CA)

On-site
USD 140,000 - 180,000
Competitive salary and equity
Full health, dental, and vision coverage
401(k) retirement savings
+4
Member of Technical Staff - ML Systems & Inference
Member of Technical Staff - ML Systems & Inference

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
Founding Engineer, ML Inference
Founding Engineer, ML Inference

Reactor • San Francisco (CA)

On-site
Competitive salary
Early equity
Health, dental, and vision coverage
+1
Senior ML Performance Engineer — Ultra-Fast Inference + Equity
Senior ML Performance Engineer — Ultra-Fast Inference + Equity

well-funded deeptech startup • California (MO)

On-site
USD 200,000 - 250,000
ML Inference Engineer San Francisco · Engineering · Full Time →
ML Inference Engineer San Francisco · Engineering · Full Time →

Reactor • San Francisco (CA)

On-site
USD 120,000 - 160,000
Visa sponsorship
Relocation support
Generous health, dental, and vision coverage
High-Performance AI Research Engineer (CUDA/ML)
High-Performance AI Research Engineer (CUDA/ML)

Metamorphic • Palo Alto (CA)

On-site
USD 200,000 - 280,000
Visa sponsorship
Competitive compensation
Mentorship and career development
+1
Founding GPU Kernel Engineer — Hand‑Tuned ML Kernels
Founding GPU Kernel Engineer — Hand‑Tuned ML Kernels

San Francisco Tensor Company • San Francisco (CA)

On-site
USD 285,000 - 315,000
Relocation assistance
Equity
Comprehensive benefits package
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 170,000 - 240,000
Staff ML Performance Engineer — Scalable Inference & CUDA
Staff ML Performance Engineer — Scalable Inference & CUDA

Modal • New York (NY)

On-site
USD 120,000 - 160,000
Senior Software Engineer - Model Performance
Senior Software Engineer - Model Performance

inference.net • San Francisco (CA)

Hybrid
USD 220,000 - 320,000
Equity in a high-growth startup
Comprehensive benefits