Get more replies from employers
Send a job-specific resume in minutes.
G-Research in London seeks an exceptional ML Performance Engineer to optimise large-scale workloads across GPU and CPU infrastructure. You will profile and tune training and inference jobs, develop reference implementations, and collaborate with research and platform teams to evolve the compute stack.
The role requires strong Python, CUDA, and PyTorch knowledge, plus experience with HPC schedulers and Kubernetes.
G-Research in London seeks an exceptional ML Performance Engineer to optimise large-scale workloads across GPU and CPU infrastructure. You will profile and tune training and inference jobs, develop reference implementations, and collaborate with research and platform teams to evolve the compute stack.
The role requires strong Python, CUDA, and PyTorch knowledge, plus experience with HPC schedulers and Kubernetes.