Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Google DeepMind seeks a Software Engineer for Model Inference to advance production-grade ML serving systems. You will collaborate with ML researchers and engineers to deploy large language models and optimize inference on GPUs/TPUs within Google's infra.
You'll work across model hosting, testing, and performance tuning, contributing to scalable, low-latency serving infrastructure and end-to-end deployment pipelines in a mission-driven team.
Google DeepMind seeks a Software Engineer for Model Inference to advance production-grade ML serving systems. You will collaborate with ML researchers and engineers to deploy large language models and optimize inference on GPUs/TPUs within Google's infra.
You'll work across model hosting, testing, and performance tuning, contributing to scalable, low-latency serving infrastructure and end-to-end deployment pipelines in a mission-driven team.