A complete application in a minute — tailored resume and cover letter, ready to send.
Kindredventures is seeking an experienced ML infra/engineering professional to design, deploy, and maintain large distributed ML training and inference clusters in a production environment.
You will build scalable pipelines for petabyte-scale data and model training, explore parallelization and numeric precision trade-offs across varying model scales, and optimize GPU performance through profiling and low-level debugging.
We value a relentless approach to problem-solving, rapid execution, and the ability to quickly learn in unfamiliar domains.
You don’t have to meet every single requirement above.