Senior Performance Co-Design Engineer, Google Cloud TPU

Google LLC

Sunnyvale (CA)

On-site

USD 163,000 - 236,000

Full time

9 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Google is seeking a Senior Performance Co-Design Engineer for Google Cloud TPU in Sunnyvale, CA. You’ll shape AI/ML hardware acceleration, drive co-design across software and hardware teams, and help push TPU architecture to the next level for demanding models.

You’ll conduct performance studies, build simulation and profiling tools, and partner with researchers to improve LLM inference latency and throughput while aligning with roadmap goals for Cloud Silicon.

Qualifications

  • Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, or a related field.
  • 8 years of experience in performance modeling/engineering, computer architecture, codesign, or systems engineering.
  • Experience programming in C++ or Python.

Responsibilities

  • Conduct comprehensive serving performance studies on current and emerging first‑party and third‑party LLMs.
  • Develop and maintain advanced simulation, profiling, and modeling tools to identify bottlenecks, understand key characteristics, and project serving workload performance.
  • Partner with model researchers, software teams, and hardware teams to co‑design architectural improvements tailored to LLM inference latency and throughput.
  • Drive data‑backed decisions that influence the roadmap for future TPU and Cloud Silicon architectures.

Skills

8 years of experience in performance
C++/Python experience

Education

Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, or related field

Tools

C++
Python

Job description

Senior Performance Co-Design Engineer, Google Cloud TPU

Share Senior Performance Co-Design Engineer, Google Cloud TPU

corporate_fare Google place Sunnyvale, CA, USA

X In most instances, this position requires in-person interviews as part of the hiring process.

  • Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, or a related field, or equivalent practical experience.
  • 8 years of experience in performance modeling/engineering, computer architecture, codesign, or systems engineering.
  • Experience programming in C++ or Python.
Preferred qualifications:
  • Master's degree or PhD in Electrical Engineering, Computer Engineering or Computer Science, with an emphasis on computer architecture.
  • Experience with hardware/software co-design problems, especially performance analysis and identification at the pre-silicon stage.
  • Experience enabling and optimizing ML models (e.g., LLMs, large embedding models).
  • Experience with ML infrastructure, profiling tools, or deep learning inference/serving optimizations.
  • Familiarity with accelerator architectures.
About the job

In this role, you’ll work to shape the future of AI/ML hardware acceleration. You will have an opportunity to drive cutting‑edge TPU (Tensor Processing Unit) technology that powers Google's most demanding AI/ML applications. You’ll be part of a team that pushes boundaries, developing custom silicon solutions that power the future of Google's TPU. You'll contribute to the innovation behind products loved by millions worldwide, and leverage your design and verification expertise to verify complex digital designs, with a specific focus on TPU architecture and its integration within AI/ML‑driven systems.

The Tensor Processing Unit (TPU) Chip Architecture and Performance Codesign team is at the forefront of optimizing Google's custom AI silicon for next‑generation machine learning models.

As a Senior Performance Co‑Design Engineer, you will focus on analyzing and optimizing the serving performance of emerging models and use cases on our custom hardware. You will also work closely with hardware architects to influence the evolution of Google’s custom ML accelerators.

The AI and Infrastructure team is redefining what’s possible. We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity. Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide.

We're the driving force behind Google's groundbreaking innovations, empowering the development of our cutting‑edge AI models, delivering unparalleled computing power to global services, and providing the essential platforms that enable developers to build the future. From software to hardware our teams are shaping the future of world‑leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.

US: $163000 - $236000 (USD) + 15% bonus target + equity + benefits

Learn more about benefits at Google .

  • Conduct comprehensive serving performance studies on current and emerging first‑party and third‑party LLMs.
  • Develop and maintain advanced simulation, profiling, and modeling tools to identify bottlenecks, understand key characteristics, and project serving workload performance.
  • Partner with model researchers, software teams, and hardware teams to co‑design architectural improvements tailored to Large Language Model (LLM) inference latency and throughput.
  • Drive data‑backed decisions that influence the roadmap for future TPU and Cloud Silicon architectures.

Google is proud to be an equal opportunity and affirmative action employer. We are committed to building a workforce that is representative of the users we serve, creating a culture of belonging, and providing an equal employment opportunity regardless of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), expecting or parents‑to‑be, criminal histories consistent with legal requirements, or any other basis protected by law. See also Google's EEO Policy , Know your rights: workplace discrimination is illegal , Belonging at Google , and How we hire .

Google is a global company and, in order to facilitate efficient collaboration and communication globally, English proficiency is a requirement for all roles unless stated otherwise in the job posting.

Equity is granted exclusively and discretionarily by Alphabet Inc. on the basis of an agreement concluded between you and Alphabet Inc. Alphabet Inc. is your sole contractual partner with respect to equity grants. GSU grants are not guaranteed, are discretionary, are subject to approval by the Alphabet Inc. board of directors or its delegate, the terms of the relevant Alphabet Inc. stock plan, and your grant agreement. They have no impact on statutory payments. Current or past grants do not confer an acquired right.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Staff Performance Codesign Engineer, TPU
Senior Staff Performance Codesign Engineer, TPU

Google LLC • Sunnyvale (CA)

On-site
USD 240,000 - 333,000
Performance Co-Design Engineer, Google Cloud TPU
Performance Co-Design Engineer, Google Cloud TPU

Google LLC • Sunnyvale (CA)

On-site
USD 192,000 - 278,000
Senior Staff Performance Codesign Engineer, TPU
Senior Staff Performance Codesign Engineer, TPU

Google • Town of Montana (WI)

On-site
USD 240,000 - 333,000
Equity
Benefits
Hardware Architecture Modeling Engineer, PhD, University Graduate, 2027 Start
Hardware Architecture Modeling Engineer, PhD, University Graduate, 2027 Start

Google Inc. • Sunnyvale (CA)

On-site
USD 138,000 - 197,000
Equity grants
Benefits
Machine Learning Hardware Architect, Google Cloud
Machine Learning Hardware Architect, Google Cloud

Google Inc. • Sunnyvale (CA)

On-site
USD 163,000 - 236,000
Software Engineer III, TPU Performance, Hardware and Software Codesign
Software Engineer III, TPU Performance, Hardware and Software Codesign

Google Inc. • Sunnyvale (CA)

On-site
USD 147,000 - 210,000
TPU Hardware Design Engineer, Cloud
TPU Hardware Design Engineer, Cloud

Google LLC • Sunnyvale (CA)

On-site
USD 138,000 - 197,000
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Google LLC • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Silicon Design Manager, RTL Machine Learning, Accelerator Design
Silicon Design Manager, RTL Machine Learning, Accelerator Design

Google LLC • Sunnyvale (CA)

On-site
USD 192,000 - 278,000
Equity
Bonus target (20%)
Benefits package
Silicon Design Manager, RTL Machine Learning, Accelerator Design
Silicon Design Manager, RTL Machine Learning, Accelerator Design

Google Inc. • Sunnyvale (CA)

On-site
USD 192,000 - 278,000
Equity grants