Senior GPU Performance Software Engineer, AI/ML

Google LLC

Greater London

In loco

GBP 90.000 - 150.000

Tempo pieno

2 giorni fa
Candidati tra i primi
Generatore di candidature

Una candidatura completa in un minuto — curriculum e lettera di presentazione personalizzati, pronti da inviare.

Supera i filtri ATS

Descrizione del lavoro

Google LLC is seeking a Senior GPU Performance Software Engineer, AI/ML, to join the ML-Omega team in London. You will help optimize, model, and evaluate GPU systems for benchmarking Google's internal ML workloads and for guiding future Cloud hardware decisions.

The role requires deep experience with ML algorithms, GPU programming, performance analysis, and collaboration with product teams to solve large-scale model performance challenges.

Competenze

  • Bachelor’s degree or equivalent practical experience.
  • 5 years of software development experience in one or more languages.
  • 5 years of data structures and algorithms experience.
  • 5 years of ML algorithms/tools, AI, DL, or NLP experience.
  • 3 years testing/maintaining/launching software products and 1 year in design/architecture.

Mansioni

  • Identify and maintain LLM benchmarks representative of Google production, industry, and the ML community, using them to identify performance opportunities, drive Accelerated Linear Algebra (XLA) GPU/Triton performance, and guide XLA releases.
  • Engage with Google product teams like DeepMind to solve ML model performance problems, onboarding new LLM models and products on GPU hardware and enabling LLMs to train and serve efficiently at a very large scale.
  • Run architecture-level simulations on GPU designs and perform roofline analysis to guide internal teams.
  • Analyze performance and efficiency metrics to identify bottlenecks, as well as design and implement solutions at Google fleet-wide scale.
  • Run performance benchmarks on GPU hardware using internal and external tools.

Conoscenze

Software development
Data structures & algorithms
Machine learning algorithms
Software design & architecture

Formazione

Bachelor's degree or equivalent practical experience

Strumenti

CUDA
Triton kernels
Palace/Mosaic kernels

Descrizione del lavoro

Senior GPU Performance Software Engineer, AI/ML

Share Senior GPU Performance Software Engineer, AI/ML

corporate_fare Google place London, UK

info_outline

info_outline X In most instances, this position requires in-person interviews as part of the hiring process.

  • Bachelor’s degree or equivalent practical experience.
  • 5 years of experience with software development in one or more programming languages.
  • 5 years of experience with data structures and algorithms.
  • 5 years of experience with machine learning algorithms and tools, artificial intelligence, deep learning, or natural language processing.
  • 3 years of experience testing, maintaining, or launching software products, and 1 year of experience with software design and architecture.
Preferred qualifications:
  • Master's degree or PhD in Computer Science or a related technical field.
  • 1 year of experience in a technical leadership role.
  • Experience with hardware/compiler co-design or high-performance computing (HPC).
  • Experience in performance analysis and debugging, improving performance of single-node or multi-node (distributed) systems.
  • Experience in GPU programming using CUDA or Triton kernels (or Palace/Mosaic kernels).
  • Background in Compiler optimizations or related fields would also be beneficial.
About the job

Google's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to Google’s needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.

About the job

Google's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to Google’s needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.

The Machine Learning (ML)-Omega team is responsible for optimizing, modeling, and evaluating GPU systems for comparative analysis and benchmarking for Google’s internal ML workloads. We strive for extracting maximum efficiency in Google’s GPU fleet, evaluating current and future ML workloads to guide decision-making for the Cloud hardware teams.

Responsibilities
  • Identify and maintain Large Language Model (LLM) training and serving benchmarks that are representative of Google production, industry, and the ML community, using them to identify performance opportunities, drive Accelerated Linear Algebra (XLA) Graphics Processing Unit (GPU)/Triton performance, and guide XLA releases.
  • Engage with Google product teams like DeepMind to solve their ML model performance problems, such as onboarding new LLM models and products on GPU hardware and enabling LLMs to train and serve efficiently at a very large scale.
  • Run architecture-level simulations on GPU designs and perform roofline analysis to guide internal teams.
  • Analyze performance and efficiency metrics to identify bottlenecks, as well as design and implement solutions at Google fleet-wide scale.
  • Run performance benchmarks on GPU hardware using internal and external tools.

Google is proud to be an equal opportunity and affirmative action employer. We are committed to building a workforce that is representative of the users we serve, creating a culture of belonging, and providing an equal employment opportunity regardless of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), expecting or parents-to-be, criminal histories consistent with legal requirements, or any other basis protected by law. See also Google's EEO Policy , Know your rights: workplace discrimination is illegal , Belonging at Google , and How we hire .

Google is a global company and, in order to facilitate efficient collaboration and communication globally, English proficiency is a requirement for all roles unless stated otherwise in the job posting.

Equity is granted exclusively and discretionarily by Alphabet Inc. on the basis of an agreement concluded between you and Alphabet Inc. Alphabet Inc. is your sole contractual partner with respect to equity grants. GSU grants are not guaranteed, are discretionary, are subject to approval by the Alphabet Inc. board of directors or its delegate, the terms of the relevant Alphabet Inc. stock plan, and your grant agreement. They have no impact on statutory payments. Current or past grants do not confer an acquired right.

Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Tech Lead, Gemini Inference Performance, DeepMind
Tech Lead, Gemini Inference Performance, DeepMind

Google • Greater London

In loco
GBP 120.000 - 180.000
Tech Lead, Gemini Inference Performance, DeepMind
Tech Lead, Gemini Inference Performance, DeepMind

Google LLC • Greater London

In loco
GBP 120.000 - 180.000
Senior GPU AI/ML Performance Engineer
Senior GPU AI/ML Performance Engineer

Google LLC • Greater London

In loco
GBP 90.000 - 150.000
Software Engineer II, Pixel Graphics and Video
Software Engineer II, Pixel Graphics and Video

Google LLC • City of Westminster

In loco
GBP 85.000 - 120.000
Software Engineer, Model Inference, DeepMind
Software Engineer, Model Inference, DeepMind

DeepMind Technologies Limited • Greater London

Ibrido
GBP 140.000 - 200.000
Tech Lead Manager, Model Performance, DeepMind
Tech Lead Manager, Model Performance, DeepMind

Google LLC • Greater London

Ibrido
GBP 198.000 - 275.000
Senior Research Engineer, ML Lead, Health Frontiers
Senior Research Engineer, ML Lead, Health Frontiers

Google • Greater London

In loco
GBP 120.000 - 180.000
Senior Research Engineer, ML Lead, Health Frontiers
Senior Research Engineer, ML Lead, Health Frontiers

Google Inc. • Greater London

In loco
GBP 90.000 - 130.000
Senior Research Engineer, ML Lead, Health Frontiers
Senior Research Engineer, ML Lead, Health Frontiers

Google Inc. • City of Westminster

In loco
GBP 120.000 - 160.000
Software Engineer, AI/ML, PhD, Early Career
Software Engineer, AI/ML, PhD, Early Career

Google Inc. • Greater London

In loco
GBP 70.000 - 90.000