Machine Learning Hardware Architect, Google Cloud

Google Inc.

Sunnyvale (CA)

On-site

USD 163,000 - 236,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Google Sunnyvale, CA seeks a Machine Learning Hardware Architect to shape TPU hardware and AI acceleration. You will drive architecture and integration across compiler, system design, and performance teams.

The role involves prototyping features, optimizing power and thermal, and collaborating with ML model teams to support training and inference at scale. Experience in C++/Python and deep learning frameworks is valued.

Qualifications

  • Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, a related field, or equivalent practical experience.
  • 5 years of experience in computer architecture, chip architecture, IP architecture, co-design, performance analysis, or hardware design.
  • Experience in developing software systems in C++ or Python.

Responsibilities

  • Create differentiated architectural innovations for Google's TPU roadmap.
  • Evaluate power, performance, and cost of prospective architecture and subsystems.
  • Collaborate with Hardware Design, Software, Compiler, ML Model and Research teams for hardware/software co-design.
  • Work on ML workload characterization and benchmarking.

Skills

Computer architecture
C++
Python
Co-design
Performance analysis
Hardware design
TensorFlow
PyTorch

Education

Bachelor's degree in EE/CE/CS or related field
Master's degree or PhD in Electrical/Computer Engineering or CS

Tools

TensorFlow
PyTorch

Job description

Machine Learning Hardware Architect, Google Cloud

Google Sunnyvale, CA, USA

  • Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, a related field, or equivalent practical experience.
  • 5 years of experience in computer architecture, chip architecture, IP architecture, co-design, performance analysis, or hardware design.
  • Experience in developing software systems in C++ or Python.
Preferred qualifications:
  • Master's degree or PhD in Electrical Engineering, Computer Engineering or Computer Science, with an emphasis on Computer Architecture, or a related field.
  • 8 years of experience in computer architecture, chip architecture, IP architecture, co-design, performance analysis, or hardware design.
  • Experience in processor design or accelerator designs and mapping ML models to hardware architectures.
  • Experience with deep learning frameworks including TensorFlow and PyTorch.
  • Knowledge of Machine Learning market, technological and business trends, software ecosystem, and emerging applications.

Knowledge of hardware/software stack for deep learning accelerators.

About the job

In this role, you'll work to shape the future of AI/ML hardware acceleration. You will have an opportunity to drive cutting-edge TPU (Tensor Processing Unit) technology that powers Google's most demanding AI/ML applications. You'll be part of a team that pushes boundaries, developing custom silicon solutions that power the future of Google's TPU. You'll contribute to the innovation behind products loved by millions worldwide, and leverage your design and verification expertise to verify complex digital designs, with a specific focus on TPU architecture and its integration within AI/ML-driven systems.

In this role, you will be at the forefront of advancing ML accelerator performance and efficiency, employing a approach that spans compiler interactions, system modeling, power architecture, and host system integration. You will prototype new hardware features, such as instruction extensions and memory layouts, by leveraging existing compiler and runtime stacks, and develop transaction-level models for early performance estimation and workload simulation. A critical part of your work will be to optimize the accelerator design for maximum performance under strict power and thermal constraints this includes evaluating novel power technologies and collaborating on thermal design. You will streamline host-accelerator interactions, minimize data transfer overheads, ensure seamless software integration across different operational modes like training and inference, and devise strategies to enhance overall ML hardware utilization. To achieve these goals, you will collaborate closely with specialized teams, including XLA (Accelerated Linear Algebra) compiler, Platforms performance, package, and system design to transition innovations to production and maintain a unified approach to modeling and system optimization.

The AI and Infrastructure team is redefining what's possible. We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity. Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide.

We're the driving team behind Google's groundbreaking innovations, empowering the development of our cutting-edge AI models, delivering unparalleled computing power to global services, and providing the essential platforms that enable developers to build the future. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $163000 - $236000 (USD) + 15% bonus target + equity + benefits

Learn more about benefits at Google .

  • Create differentiated architectural innovations for Google's semiconductor Tensor Processing Unit (TPU) roadmap.
  • Evaluate the power, performance, and cost of prospective architecture and subsystems.
  • Collaborate with partners in Hardware Design, Software, Compiler, ML Model and Research teams for hardware/software co-design.
  • Work on Machine Learning (ML) workload characterization and benchmarking.
  • Develop architecture for differentiating features on next generation TPUs.

Google is proud to be an equal opportunity and affirmative action employer. We are committed to building a workforce that is representative of the users we serve, creating a culture of belonging, and providing an equal employment opportunity regardless of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), expecting or parents-to-be, criminal histories consistent with legal requirements, or any other basis protected by law. See also Google's EEO Policy , Know your rights: workplace discrimination is illegal , Belonging at Google , and How we hire .

Google is a global company and, in order to facilitate efficient collaboration and communication globally, English proficiency is a requirement for all roles unless stated otherwise in the job posting.

To all recruitment agencies: Google does not accept agency resumes. Please do not forward resumes to our jobs alias, Google employees, or any other organization location. Google is not responsible for any fees related to unsolicited resumes.

Equity is granted exclusively and discretionarily by Alphabet Inc. on the basis of an agreement concluded between you and Alphabet Inc. Alphabet Inc. is your sole contractual partner with respect to equity grants. GSU grants are not guaranteed, are discretionary, are subject to approval by the Alphabet Inc. board of directors or its delegate, the terms of the relevant Alphabet Inc. stock plan, and your grant agreement. They have no impact on statutory payments. Current or past grants do not confer an acquired right.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Hardware Architect, Google Cloud
Machine Learning Hardware Architect, Google Cloud

Google • Sunnyvale (CA)

On-site
USD 163,000 - 236,000
Senior Staff Performance Codesign Engineer, TPU
Senior Staff Performance Codesign Engineer, TPU

Google • Town of Montana (WI)

On-site
USD 240,000 - 333,000
Equity
Benefits
Hardware Architecture Modeling Engineer, PhD, University Graduate, 2027 Start
Hardware Architecture Modeling Engineer, PhD, University Graduate, 2027 Start

Google Inc. • Sunnyvale (CA)

On-site
USD 138,000 - 197,000
Equity grants
Benefits
Senior RTL Design Engineer, TPU Compute
Senior RTL Design Engineer, TPU Compute

Google • Sunnyvale (CA)

On-site
USD 163,000 - 236,000
TPU SoC Design Engineer, Google Cloud
TPU SoC Design Engineer, Google Cloud

Google • Town of Montana (WI)

On-site
USD 116,000 - 165,000
Hardware Architecture Modeling Engineer, PhD, University Graduate, 2027 Start
Hardware Architecture Modeling Engineer, PhD, University Graduate, 2027 Start

Google • Sunnyvale (CA)

On-site
USD 138,000 - 197,000
Equity
Bonus target
Benefits
Software Engineer III, TPU Performance, Hardware and Software Codesign
Software Engineer III, TPU Performance, Hardware and Software Codesign

Google Inc. • Sunnyvale (CA)

On-site
USD 147,000 - 210,000
Senior Staff Performance Codesign Engineer, TPU
Senior Staff Performance Codesign Engineer, TPU

Google Inc. • Sunnyvale (CA)

On-site
USD 240,000 - 333,000
Bonus target
Equity
Benefits
RTL Design Engineer, Machine Learning Accelerators
RTL Design Engineer, Machine Learning Accelerators

Google Inc. • Sunnyvale (CA)

On-site
USD 138,000 - 197,000
Equity
Bonus target
Senior Software Engineer, Fleet-level ML Performance
Senior Software Engineer, Fleet-level ML Performance

Google • United States

On-site
USD 174,000 - 252,000