Senior ML Performance Engineer — Ultra-Fast Inference + Equity

well-funded deeptech startup

California (MO)

On-site

USD 200,000 - 250,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Join a well-funded deeptech startup in California as a Senior ML Performance Engineer. You will lead the optimization of ML inference pipelines and improve model performance for innovative clients.

This role requires strong skills in Python, TensorFlow, and PyTorch, as well as expertise in performance optimization techniques. The ideal candidate is eager to work on cutting-edge ML projects in an energetic environment with potential for significant impact.

Qualifications

  • Strong proficiency in Python and ML frameworks such as TensorFlow and PyTorch.
  • Deep understanding of ML algorithms and architectures.
  • Expertise in performance optimization techniques including profiling and quantization.

Responsibilities

  • Conduct in-depth performance profiling and analysis of ML models.
  • Design and implement efficient ML inference pipelines.
  • Collaborate with infrastructure teams to optimize configurations.

Skills

Python
TensorFlow
PyTorch
Performance optimization techniques
Kubernetes
AWS
GCP
Azure

Job description

Join a well-funded deeptech startup in California as a Senior ML Performance Engineer. You will lead the optimization of ML inference pipelines and improve model performance for innovative clients.

This role requires strong skills in Python, TensorFlow, and PyTorch, as well as expertise in performance optimization techniques. The ideal candidate is eager to work on cutting-edge ML projects in an energetic environment with potential for significant impact.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Performance Engineer
Senior ML Performance Engineer

well-funded deeptech startup • California (MO)

On-site
USD 200,000 - 250,000
ML Inference Engineer San Francisco · Engineering · Full Time →
ML Inference Engineer San Francisco · Engineering · Full Time →

Reactor • San Francisco (CA)

On-site
USD 120,000 - 160,000
Visa sponsorship
Relocation support
Generous health, dental, and vision coverage
High-Performance ML Inference Engineer
High-Performance ML Inference Engineer

Reactor • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive SF salary
Early equity
Visa sponsorship
+2
Founding ML Inference Performance Engineer
Founding ML Inference Performance Engineer

uRun • San Francisco (CA)

On-site
USD 120,000 - 160,000
Health, dental, and vision
401(k) participation
Flexible spending accounts
+3
Software Engineer - ML Model Performance
Software Engineer - ML Model Performance

Baseten • San Francisco (CA)

On-site
USD 150,000 - 250,000
Senior ML Performance Engineer: Scale & Throughput
Senior ML Performance Engineer: Scale & Throughput

NLP PEOPLE • Sunnyvale (CA)

On-site
USD 215,000 - 285,000
ML Performance Engineer – Real-Time Inference
ML Performance Engineer – Real-Time Inference

Odyssey • Palo Alto (CA)

On-site
USD 130,000 - 160,000
Senior ML Engineer — Ultra-Fast Inference & Memory
Senior ML Engineer — Ultra-Fast Inference & Memory

comfy-org • San Francisco (CA)

On-site
USD 100,000 - 140,000
Senior ML Inference Engineer — AI Infrastructure & Equity
Senior ML Inference Engineer — AI Infrastructure & Equity

Oscar Technology • San Francisco (CA)

Hybrid
USD 225,000 - 275,000
Equity
401k matching
Medical coverage
+1
Machine Learning Inference Engineer
Machine Learning Inference Engineer

Oscar Technology • San Francisco (CA)

Hybrid
USD 225,000 - 275,000
Equity
401k matching
Medical coverage
+1