Staff GenAI Inference Performance Engineer

Google

Mountain View (CA)

On-site

USD 207,000 - 300,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Bonus target
Benefits

Job summary

Google DeepMind is seeking a Staff Software Engineer focused on Inference Performance Optimization for GenAI. Based in Mountain View, you will drive optimization of AI inference workloads, design fast serving techniques, and analyze bottlenecks to maximize throughput and minimize latency across distributed systems.

You will work with a mission-driven team advancing AI agents, with emphasis on scalable inference, profiling, and performance engineering.

Qualifications

  • Bachelor's degree or equivalent practical experience in CS/CE/Applied Math or related field.
  • 8 years of software development experience.
  • Experience in Python and C++, including navigating, debugging, and modifying serving codebases.
  • Experience with AI model execution constraints and modern serving architectures.

Responsibilities

  • Analyze and optimize AI inference workloads to increase throughput per GPU and reduce latency.
  • Design and implement inference optimization techniques.
  • Investigate and resolve complex model inference performance bottlenecks across the stack.
  • Model latency-to-cost impacts and translate insights into actionable production signals.
  • Develop investigative tools and metrics to track compute usage across the fleet.

Skills

Python
C++
Debugging
Serving architectures

Education

Bachelor's degree or equivalent experience

Job description

Google DeepMind is seeking a Staff Software Engineer focused on Inference Performance Optimization for GenAI. Based in Mountain View, you will drive optimization of AI inference workloads, design fast serving techniques, and analyze bottlenecks to maximize throughput and minimize latency across distributed systems.

You will work with a mission-driven team advancing AI agents, with emphasis on scalable inference, profiling, and performance engineering.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer, Inference Performance Optimization, GenAI, DeepMind
Staff Software Engineer, Inference Performance Optimization, GenAI, DeepMind

Google • Mountain View (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
AI Inference Compute Engineer — Hardware/Software Co-Design
AI Inference Compute Engineer — Hardware/Software Co-Design

Google DeepMind • San Francisco (CA)

On-site
USD 174,000 - 252,000
Staff GenAI Inference Engineer: Optimize LLM Serving Latency
Staff GenAI Inference Engineer: Optimize LLM Serving Latency

Menlo Ventures • San Francisco (CA)

On-site
USD 190,000 - 233,000
Annual performance bonus
Equity options
Comprehensive health benefits
Staff GenAI Evaluation Engineer
Staff GenAI Evaluation Engineer

Google Inc. • Sunnyvale (CA)

On-site
USD 207,000 - 301,000
Staff GenAI Kernel & Performance Engineer
Staff GenAI Kernel & Performance Engineer

Databricks • San Francisco (CA)

On-site
USD 190,000 - 233,000
Distributed AI Inference Performance Engineer
Distributed AI Inference Performance Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Staff GenAI Inference Architect: High-Throughput ML Serving
Staff GenAI Inference Architect: High-Throughput ML Serving

Databricks • San Francisco (CA)

On-site
USD 190,000 - 233,000
Senior Performance AI & Infra Engineer
Senior Performance AI & Infra Engineer

Google • Kirkland (WA)

On-site
USD 174,000 - 253,000
Health, dental, vision, life, and disb
401(k) with company match
Paid time off 20 days/year
+4
Senior Staff AI/ML GenAI Engineer — Cloud Scale
Senior Staff AI/ML GenAI Engineer — Cloud Scale

Google • Sunnyvale (CA)

On-site
USD 248,000 - 349,000
Senior AI Inference Performance Engineer — Scale GPUs
Senior AI Inference Performance Engineer — Scale GPUs

NVIDIA • California (MO)

On-site
USD 124,000 - 196,000
Equity eligibility