A cutting-edge AI infrastructure company in San Francisco is seeking a Senior Technical Leader to design and lead the development of a next-generation model serving platform. This role demands strong skills in C++ and Python, along with considerable experience in building scalable systems. The successful candidate will mentor engineers and ensure high-performance execution while collaborating closely with ML researchers. The role offers competitive compensation, including salary and equity, while fostering a high-ownership culture.
Qualifications
5+ years of experience designing and building scalable backend systems.
Strong understanding of LLM inference mechanics.
Experience with containerization and distributed infrastructure.
Responsibilities
Lead the technical direction of the model serving platform.
Build core serving components including execution runtimes and distributed inference systems.
Develop high-performance C++ and CUDA/HIP modules.
Skills
C++
Python
Debugging
Performance optimization
Collaboration with ML researchers
Leadership skills
Education
Bachelor’s degree in Computer Science or equivalent
Tools
Kubernetes
CUDA
HIP
Job description
A cutting-edge AI infrastructure company in San Francisco is seeking a Senior Technical Leader to design and lead the development of a next-generation model serving platform. This role demands strong skills in C++ and Python, along with considerable experience in building scalable systems. The successful candidate will mentor engineers and ensure high-performance execution while collaborating closely with ML researchers. The role offers competitive compensation, including salary and equity, while fostering a high-ownership culture.