Founding ML Inference Engineer — Ultra-Low Latency AI

Reactor

San Francisco (CA)

On-site

USD 180,000 - 280,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive salary
Early equity
Health, dental, and vision coverage
Relocation support

Job summary

A media technology company in San Francisco is seeking a Founding Engineer specializing in ML Inference. This highly technical role requires expertise in the ML infrastructure stack and aims to optimize generative media performance. The ideal candidate will drive innovations in real-time model performance, design in-house inference runtimes, and optimize models through advanced techniques. Competitive salary and relocation support are offered, along with generous health coverage.

Qualifications

  • Strong foundation in systems programming with a track record of identifying and resolving bottlenecks.
  • Deep expertise in ML infrastructure stack, including PyTorch, TensorRT, and more.
  • Knowledge of model compilation and quantization techniques.

Responsibilities

  • Drive performance for real-time model performance for diffusion models.
  • Design a high-performance in-house inference runtime.
  • Optimize models for inference through quantization and pruning.
  • Benchmark model performance to identify bottlenecks.

Skills

Systems programming
PyTorch
TensorRT
Model compilation
Quantization
NVIDIA GPU hardware knowledge
Transformer architectures

Job description

A media technology company in San Francisco is seeking a Founding Engineer specializing in ML Inference. This highly technical role requires expertise in the ML infrastructure stack and aims to optimize generative media performance. The ideal candidate will drive innovations in real-time model performance, design in-house inference runtimes, and optimize models through advanced techniques. Competitive salary and relocation support are offered, along with generous health coverage.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Founding Engineer, ML Inference
Founding Engineer, ML Inference

Reactor • San Francisco (CA)

On-site
USD 180,000 - 280,000
Competitive salary
Early equity
Health, dental, and vision coverage
+1
Staff ML Engineer — Ultra-Low-Latency Inference
Staff ML Engineer — Ultra-Low-Latency Inference

Inworld • Mountain View (CA)

Hybrid
USD 270,000 - 500,000
Relocation assistance
Equity options
Comprehensive benefits
High-Performance ML Inference Engineer
High-Performance ML Inference Engineer

Reactor • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive SF salary
Early equity
Visa sponsorship
+2
Member of Technical Staff, Inference
Member of Technical Staff, Inference

Reactor • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive SF salary
Early equity
Visa sponsorship
+2
ML Inference Engineer San Francisco · Engineering · Full Time →
ML Inference Engineer San Francisco · Engineering · Full Time →

Reactor • San Francisco (CA)

On-site
USD 120,000 - 160,000
Visa sponsorship
Relocation support
Generous health, dental, and vision coverage
Founding ML Inference Performance Engineer
Founding ML Inference Performance Engineer

uRun • San Francisco (CA)

On-site
USD 120,000 - 160,000
Health, dental, and vision
401(k) participation
Flexible spending accounts
+3
Inference Optimization Engineer: Fast, Cost-Effective ML
Inference Optimization Engineer: Fast, Cost-Effective ML

Build AI • San Francisco (CA)

On-site
USD 150,000 - 210,000
Competitive pay
Medical, dental, and vision packages
Housing subsidy $2k/month near SF offi
+6
ML Inference Infrastructure Engineer
ML Inference Infrastructure Engineer

Baseten • San Francisco (CA)

On-site
USD 90,000 - 130,000
100% coverage of medical, dental, and vision insurance
Generous PTO policy including Winter Break
Company-facilitated 401(k)
ML Inference Systems Engineer
ML Inference Systems Engineer

Gimlet Labs, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
AI Inference Engineer: Real-Time ML, Hybrid, Equity
AI Inference Engineer: Real-Time ML, Hybrid, Equity

Pantera Capital • Palo Alto (CA)

Hybrid
USD 190,000 - 250,000
Comprehensive health insurance
Dental insurance
Vision insurance
+1