Staff ML Inference Engineer — Model Efficiency (Remote)

Jaide Health

San Francisco (CA)

Presencial

USD 120.000 - 160.000

Jornada completa

14 días+
Generador de candidaturas

Convierte este puesto en una entrevista — un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Inclusive culture and work environment
Weekly lunch stipend, in-office lunches & snacks
Full health and dental benefits
100% Parental Leave top-up for up to 6 months
6 weeks of vacation (30 working days!)

Descripción de la vacante

Jaide Health is seeking an engineer for their Model Efficiency team in San Francisco. The role focuses on building reliable ML systems while enhancing core performance metrics across model execution. You'll work with advanced performance techniques such as GPU/CUDA optimizations and collaborate closely with modeling and systems teams. Ideal candidates will have over 5 years of experience in high-performance coding, plus strong skills in C++ or Python and insights into the LLM inference ecosystem. A commitment to diversity and inclusive work culture is celebrated.

Formación

  • 5+ years of experience writing high-performance, production-quality code.
  • Strong programming skills in C++ or Python (Rust/Go also welcome).
  • Experience working with large language models and familiarity with the LLM inference ecosystem.

Responsabilidades

  • Work across the inference stack to improve core performance metrics.
  • Identify bottlenecks and develop optimizations for model execution.
  • Collaborate closely with modeling and systems teams to measure and ship improvements.

Conocimientos

High-performance, production-quality code
Programming in C++ or Python
Diagnosing and resolving performance bottlenecks
GPU programming
Experience with large language models

Descripción del empleo

Jaide Health is seeking an engineer for their Model Efficiency team in San Francisco. The role focuses on building reliable ML systems while enhancing core performance metrics across model execution. You'll work with advanced performance techniques such as GPU/CUDA optimizations and collaborate closely with modeling and systems teams. Ideal candidates will have over 5 years of experience in high-performance coding, plus strong skills in C++ or Python and insights into the LLM inference ecosystem. A commitment to diversity and inclusive work culture is celebrated.
Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Staff Engineer - ML Inference & Model Efficiency
Staff Engineer - ML Inference & Model Efficiency

Cohere • San Francisco (CA)

A distancia
USD 180.000 - 240.000
Inclusive work culture
Weekly lunch stipend
Full health and dental benefits
+4
Senior ML Systems Engineer - Model Inference & Efficiency
Senior ML Systems Engineer - Model Inference & Efficiency

Cohere • New York (NY)

Híbrido
USD 100.000 - 150.000
Inclusive culture and work environment
Weekly lunch stipend, in-office lunches & snacks
Full health and dental benefits
+4
Software Engineer - ML Model Performance
Software Engineer - ML Model Performance

Baseten • San Francisco (CA)

Presencial
USD 150.000 - 250.000
Competitive compensation with equity
100% medical, dental, and vision insurance
Generous PTO policy
+2
Senior Staff Engineer, Model Efficiency - Remote
Senior Staff Engineer, Model Efficiency - Remote

Cohere • EE. UU.

A distancia
USD 150.000 - 210.000
ML Model Performance Engineer - Inference and Acceleration
ML Model Performance Engineer - Inference and Acceleration

Baseten • New York (NY)

Presencial
USD 200.000 - 275.000
100% coverage of medical, dental, and vision insurance
Generous PTO policy
Paid parental leave
+2
Senior ML Inference Systems Engineer - Remote-Friendly
Senior ML Inference Systems Engineer - Remote-Friendly

Atlassian • Mountain View (CA)

Híbrido
USD 206.000 - 269.000
Health & wellbeing
Volunteer days
Community perks
Staff SWE, Inference Infrastructure — High-Scale ML
Staff SWE, Inference Infrastructure — High-Scale ML

Jaide Health • San Francisco (CA)

Híbrido
USD 130.000 - 170.000
Open and inclusive culture
Weekly lunch stipend and snacks
Full health and dental benefits
+3
Staff ML Engineer - Frontier AI for Clinical Excellence
Staff ML Engineer - Frontier AI for Clinical Excellence

Ambience Healthcare • San Francisco (CA)

Híbrido
USD 250.000 - 350.000
Comprehensive medical, dental and vision coverage
401(k) with company match
Flexible time off with no cap
Remote Audio Inference Engineer, Model Efficiency
Remote Audio Inference Engineer, Model Efficiency

Jaide Health • San Francisco (CA)

Presencial
USD 100.000 - 140.000
Open and inclusive culture
Weekly lunch stipend
Full health and dental benefits
+2
ML Inference Systems Engineer
ML Inference Systems Engineer

Gimlet Labs, Inc. • San Francisco (CA)

Presencial
USD 120.000 - 160.000