Principal Engineer, Efficient GenAI

AMD

San Jose (CA)

On-site

USD 250,000 - 360,000

Full time

8 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

AMD in San Jose, CA is seeking a Principal Engineer to advance Generative AI training and inference at scale. You will join the AI Models and Applications team to design efficient transformer architectures and scalable distributed training on AMD hardware.

Qualifications include a PhD or master's in CS/EE/math, deep experience with LLMs or image/video models, and fluency with PyTorch, JAX, vLLM, SGLang, or MuJoCo.

Qualifications

  • You have a deep technical understanding of Generative AI applications (LLMs, 3D World/Action models, or image/video generation).
  • You have experience training models at scale and enabling distributed training and inference on AMD devices.
  • You hold a PhD or master's in a relevant field and have several years in AI/deep learning.

Responsibilities

  • Propose and apply innovative techniques for both training and inference, including transformer architectures and parallelism strategies.
  • Develop novel architectures for Generative AI models and showcase benefits on AMD platforms.
  • Collaborate with open-source communities to integrate AMD-optimised models and publish training recipes.
  • Co‑optimise end‑to‑end performance with software and hardware teams; promote scalable AI workflows.

Skills

Generative AI expertise
Distributed training
Technical leadership

Education

PhD or master's in CS/EE/Math

Tools

PyTorch
JAX
vLLM
SGLang
MuJoCo

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.

Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger— technologythat moves the world forward. Join us and, together, we’ll advance your career.

THE ROLE:

The AI Models and Applications team at AMD is looking for a specialized Principal Engineer who is passionate about enabling innovative and efficient Generative AI training and inference at scale. You will be part of a core team of incredibly talented specialists and work on scaling training and inference for the latest Generative AI models.

THE PERSON:

You have a deep technical understanding and hands‑on experience with the latest Generative AI applications in at least one of the following areas: large language models (LLMs), 3D World and Action Models, or image/video generation models. You have experience training models at scale and are passionate about developing efficient approaches to enable distributed training and inference on AMD devices.

Why Join Us?
  • Exciting Opportunities: As a senior member of the team, you will be at the forefront of innovation, working with the latest Generative AI models and algorithms. You will have the opportunity to shape the future of AI model training and inference optimisation across a variety of applications.
  • Talented Team: Join a team of highly skilled industry specialists who are passionate about pushing the boundaries of AI. Collaborate with like‑minded professionals and learn from the best in the field.
  • Cutting‑Edge Technology: Work with state‑of‑the‑art Generative AI algorithms and software, enabling you to stay ahead of the curve and drive advancements in AI model training at scale and deployment.
  • Impactful Work: Your contributions will directly influence how cutting‑edge Generative AI models across the industry are efficiently trained at scale, as well as how inference solutions are deployed to serve millions of customers, making a significant difference across various industries and applications.
KEY RESPONSIBILITIES:
  • Propose and apply innovative techniques to support both training and inference, including innovative transformer architectures, parallelism strategies for training on large clusters, inference optimisation techniques such as speculative decoding, and optimal KV‑caching strategies.
  • Implement novel, efficient architectures for Generative AI models for training and inference and showcase the benefits on AMD platforms.
  • Work with open‑source frameworks and communities (e.g., PyTorch, JAX, vLLM, SGLang) to integrate AMD‑optimised models and libraries and publish training recipes.
  • Collaborate with software and hardware teams to co‑optimise end‑to‑end performance on current and future AMD solutions.
  • Increase adoption of agentic workflows for optimising and deploying Generative AI applications at scale on AMD platforms.
  • Publish and promote your work at external venues, including major conferences.
  • Collaborate with researchers within AMD and across industry and academia to promote innovation on AMD platforms.
PREFERRED EXPERIENCE:
  • Strong technical expertise in Generative AI model training and inference, with familiarity working with deep learning frameworks such as PyTorch, JAX, vLLM, SGLang, and MuJoCo.
  • Strong technical expertise in algorithmic innovation for efficient Generative AI applications across both training and inference.
  • Expertise and publications in one or more of the following preferred areas: efficient model architectures, optimised training, innovative parallelism strategies, or low‑precision training.
  • Additional plus if publications have been presented at conferences such as NeurIPS, CVPR, ECCV, ICCV, ICML, or ICLR.
  • Experience productising Generative AI models and training foundation models at scale.
  • Excellent written, verbal, and presentation skills, with the ability to coordinate effectively both internally and externally.
  • Several years of experience in AI, deep learning, and related software development.
ACADEMIC CREDENTIALS:

PhD or master’s degree in computer science, Electrical Engineering, Mathematics, or a related field.

LOCATION:

San Jose, CA (Hybrid)

Alternative locations in Seattle, WA, or Austin, TX may be considered.

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee‑based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third‑party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Principal Engineer, Efficient GenAI
Principal Engineer, Efficient GenAI

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 180,000 - 250,000
Principal Software Engineer — AI Performance & Reliability
Principal Software Engineer — AI Performance & Reliability

Advanced Micro Devices, Inc. • San Jose (CA)

Hybrid
USD 250,000 - 420,000
Hybrid work model
Benefits package
Principal GenAI Inference Optimization Engineer
Principal GenAI Inference Optimization Engineer

Advanced Micro Devices • San Jose (CA)

On-site
USD 150,000 - 200,000
Frontier AI Workloads - Performance and Scalability Engineer
Frontier AI Workloads - Performance and Scalability Engineer

AMD • San Jose (CA)

On-site
USD 150,000 - 200,000
Competitive salary
Health benefits
Career advancement opportunities
Fellow Software Engineer - AI Performance & Reliability
Fellow Software Engineer - AI Performance & Reliability

Advanced Micro Devices • San Jose (CA)

On-site
USD 180,000 - 240,000
Fellow Software Engineer — AI Performance & Reliability
Fellow Software Engineer — AI Performance & Reliability

AMD • San Jose (CA)

On-site
USD 180,000 - 260,000
AI Research Scientist
AI Research Scientist

Advanced Micro Devices, Inc. • Bellevue (WA)

On-site
USD 180,000 - 280,000
Benefits at a glance
Fellow GPU Performance Optimization Engineer
Fellow GPU Performance Optimization Engineer

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 150,000 - 180,000
Fellow GPU Performance Engineer AI Training at Scale
Fellow GPU Performance Engineer AI Training at Scale

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
Confidential
AI/ML Solutions Engineer
AI/ML Solutions Engineer

Advanced Micro Devices, Inc. • Santa Clara (CA)

On-site
USD 150,000 - 230,000