Get more replies from employers
Send a job-specific resume in minutes.
Generalist is seeking a talented individual in Somerville, Massachusetts to manage GPU fleets for large-scale robot training. The role involves optimizing data loading systems and orchestrating inference fleets, ensuring high performance under compute constraints.
The ideal candidate has deep experience with Slurm or Kubernetes for machine learning orchestration and a solid understanding of the Nvidia GPU ecosystem. Come join an equal opportunity employer that values diversity!
Generalist trains very large robot foundation models. This requires utilizing very large numbers of the latest generation GPU hardware and infrastructure (currently Nvidia) to run distributed training jobs and researcher experiments. We have extreme requirements on storage and data loading infrastructure that requires maximizing cloud infrastructure and custom solutions.
You will also own inference infrastructure. For our robots this is a fleet of on-prem GPUs attached to robots that have extreme real-time and latency budgets in compute constrained environments.
You’ll be responsible for:We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.