Staff ML Performance Engineer: Training Efficiency & Scale
Icehouseventures
Sunnyvale (CA)
On-site
USD 120,000 - 160,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
Wayve is seeking a Staff ML Performance Engineer to join their Training Tech team in Sunnyvale, California. This role focuses on optimizing large scale ML jobs to enhance training and inference workloads. The ideal candidate will have over 10 years of industry experience in performance engineering, proficiency in GPU optimization, and the ability to write high-quality Python code. You'll collaborate with research teams to drive performance improvements, contributing to the advancement of self-driving technology.
Qualifications
10+ years of industry experience in performance engineering across ML systems.
Experience optimizing large scale jobs on GPU compute clusters.
Ability to write high quality, well-structured and tested Python code.
Responsibilities
Profile ML workloads to identify their bottlenecks.
Design and implement efficiency improvements.
Collaborate closely with research teams for training efficiency.
Skills
Performance engineering
GPU compute optimization
Research team collaboration
Python programming
Education
BS or MS in Machine Learning, Computer Science, Engineering or related technical discipline
Tools
NVIDIA NSight Systems
CUDA
Job description
Wayve is seeking a Staff ML Performance Engineer to join their Training Tech team in Sunnyvale, California. This role focuses on optimizing large scale ML jobs to enhance training and inference workloads. The ideal candidate will have over 10 years of industry experience in performance engineering, proficiency in GPU optimization, and the ability to write high-quality Python code. You'll collaborate with research teams to drive performance improvements, contributing to the advancement of self-driving technology.