ML Systems Engineer: LLM Deployment & Performance

DeepMind Technologies Limited

Madison (WI)

On-site

USD 207,000 - 300,000

Full time

37 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Google DeepMind is seeking a Software Engineer to help deploy and optimize large language models and AI agents on our production infrastructure. You will work with researchers and engineers to build efficient serving, testing, and validation pipelines, and contribute to performance tuning across hardware accelerators.

You will have opportunities across IC and TL tracks, with roles spanning software and research engineering, and you will contribute to a high-impact, mission-driven team focused on

Qualifications

  • Bachelor’s degree or equivalent practical experience.
  • 8 years of experience in software development.
  • 2 years of experience deploying and maintaining ML models in production.
  • Experience profiling, configuring, or executing ML workloads on hardware accelerators (GPU/TPU).
  • Experience designing, building, or optimizing model serving infrastructure.

Responsibilities

  • Collaborate with Research teams to understand next generation modeling approaches with production considerations.
  • Work with infrastructure teams to deliver serving infrastructure for speed, scale and quality.
  • Identify opportunities to automate tasks, reduce redundancies, and improve model release velocity.
  • Gain understanding of serving frameworks, pre-processing pipelines, caching, and related tech.
  • Use roofline analysis and profiling to find and fix performance bottlenecks across ML frameworks, compilers (XLA), kernels (Pallas), and serving infra on TPUs/GPUs.

Skills

Software development
ML model deployment
ML workloads profiling
Model serving infrastructure
Hardware accelerators (GPUs/TPUs)

Education

Bachelor’s degree or equivalent practical experience

Tools

GPUs
TPUs

Job description

Google DeepMind is seeking a Software Engineer to help deploy and optimize large language models and AI agents on our production infrastructure. You will work with researchers and engineers to build efficient serving, testing, and validation pipelines, and contribute to performance tuning across hardware accelerators.

You will have opportunities across IC and TL tracks, with roles spanning software and research engineering, and you will contribute to a high-impact, mission-driven team focused on

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Inference Engineer - LLM Deployment & Serving
ML Inference Engineer - LLM Deployment & Serving

Google DeepMind • Mountain View (CA)

Hybrid
USD 230,000 - 290,000
Software Engineer, Model Inference & LLM Deployment
Software Engineer, Model Inference & LLM Deployment

Google LLC • Mountain View (CA), Northern (KY)

Hybrid
USD 180,000 - 300,000
ML Systems Engineer — LLM Inference & Multi-Node Performance
ML Systems Engineer — LLM Inference & Multi-Node Performance

ScOp Venture Capital LLC. • Santa Barbara (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Competitive compensation
Equity
Professional growth opportunities
+1
Software Engineer, Model Inference, DeepMind
Software Engineer, Model Inference, DeepMind

Google DeepMind • Mountain View (CA)

Hybrid
USD 230,000 - 290,000
ML Systems Engineer
ML Systems Engineer

ScOp Venture Capital LLC. • Santa Barbara (CA), Northern (KY)

On-site
USD 150,000 - 210,000
Competitive compensation
Equity
Professional growth opportunities
+1
Senior ML Engineer – AI/NLP & RL Systems
Senior ML Engineer – AI/NLP & RL Systems

Google • Mountain View (CA)

On-site
USD 174,000 - 252,000
Bonus target
Equity grant
Benefits
GenAI Research Engineer — LLM & Model Evaluation
GenAI Research Engineer — LLM & Model Evaluation

Google Inc. • Cambridge (MA), Northern (KY)

Hybrid
USD 174,000 - 252,000
ML Engineer — Production-Grade AI & LLM/VLM Systems
ML Engineer — Production-Grade AI & LLM/VLM Systems

Nace.AI • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Senior Staff Research Scientist — LLMs & Post-Training AI
Senior Staff Research Scientist — LLMs & Post-Training AI

DeepMind Technologies Limited • Kirkland (WA)

On-site
USD 207,000 - 300,000
Senior AI/ML Research Engineer - LLMs & Production Systems
Senior AI/ML Research Engineer - LLMs & Production Systems

Google • Mountain View (CA)

On-site
USD 174,000 - 252,000