A pioneering AI firm is seeking a Staff Software Engineer for ML Acceleration to enhance the product development velocity related to machine learning. Responsibilities include analyzing ML models to optimize performance, integrating tools for ML engineers, and implementing solutions across hardware platforms. The ideal candidate should possess strong programming skills in C++ and Python, have over 5 years of experience, and familiarity with deep learning frameworks like PyTorch. A commitment to engineering excellence is essential in this role, which may require remote work options.
Qualifications
5+ years of experience with strong focus on GPU programming and optimization.
Experience with deep learning frameworks, particularly PyTorch.
Excellent verbal and written communication skills.
Responsibilities
Analyze ML models to identify and resolve performance bottlenecks.
Incorporate OSS tools for ML engineers to profile and optimize models.
Implement optimizations using CUDA, Triton, and custom kernels.
Skills
C++ programming
Python programming
GPU programming
Deep learning frameworks (PyTorch)
CUDA programming
Triton language
TensorRT implementation
ONNX model conversion
Analytical skills
Communication skills
Education
Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field
Tools
CUDA
Triton
PyTorch
Job description
A pioneering AI firm is seeking a Staff Software Engineer for ML Acceleration to enhance the product development velocity related to machine learning. Responsibilities include analyzing ML models to optimize performance, integrating tools for ML engineers, and implementing solutions across hardware platforms. The ideal candidate should possess strong programming skills in C++ and Python, have over 5 years of experience, and familiarity with deep learning frameworks like PyTorch. A commitment to engineering excellence is essential in this role, which may require remote work options.