Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Cerence Inc. is seeking an ML Systems Engineer to design and operate distributed training infrastructure for large neural networks across GPU clusters. You will optimize multi-node, multi-GPU execution and diagnose bottlenecks in compute, memory, and networking.
You will collaborate with research and applied ML teams to productionize training pipelines, manage multi-node launches, handle failure recovery, and tune performance for reliability at scale.
Cerence Inc. is seeking an ML Systems Engineer to design and operate distributed training infrastructure for large neural networks across GPU clusters. You will optimize multi-node, multi-GPU execution and diagnose bottlenecks in compute, memory, and networking.
You will collaborate with research and applied ML teams to productionize training pipelines, manage multi-node launches, handle failure recovery, and tune performance for reliability at scale.