Fellow, AI Performance Software Engineer

Advanced Micro Devices, Inc.

Santa Clara (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Advanced Micro Devices, Inc. is seeking talented AI Software Engineers to join their cutting-edge AI Software Solutions Team in Santa Clara, California. You will be involved in optimizing the software ecosystem for next-generation GPU computational accelerators, requiring a minimum of 4 years of related experience.

The ideal candidate will have strong programming skills in C++ and Python, with experience in deep learning frameworks. Candidates should be prepared to work collaboratively in a dynamic environment focused on innovation and excellence.

Qualifications

  • Minimum 4 years of experience required.
  • Strong analytical and problem-solving skills.
  • Experience with open-source software development is a plus.

Responsibilities

  • Enable DL models for Instinct GPUs in cloud and on-premise environments.
  • Analyze and optimize the performance of AI software.
  • Collaborate with Software Engineers on performance improvements.

Skills

C++ programming
Python programming
Deep Learning frameworks (PyTorch/TensorFlow)
Performance optimization

Education

M.S. or Ph.D. in Computer Science/Engineering

Tools

TorchProfiler
RocM profiler
VTune
Nsight
Docker
Kubernetes

Job description

WHAT YOU DO AT AMD CHANGES EVERYTHING

At AMD, our mission is to build great products that accelerate next‑generation computing experiences – from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges, striving for execution excellence while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond.

Together, we advance your career.

THE GROUP

AI is defining the next era of computing, and this is just the beginning. We see the benefits of AI every day – enabling medical research, curbing credit card fraud, reducing congestion in cities, or simply making life easier. In the ever‑evolving landscape of artificial intelligence, we are a powerhouse – a cutting‑edge AI Software Solutions Team. Specialized in AI optimisation, fine‑tuning large language models to unlock unprecedented Generative AI efficiency, our expertise extends beyond the hardware realm into 3P enablement, where we develop custom AI software solutions for industry‑leading AI customers.

THE ROLE

We are searching for talented and highly motivated AI Software Engineers to join our team of developers pushing the boundaries of efficiency and performance to enable and optimise the software ecosystem for the next generation of GPU computational accelerators. Minimum 4 years of experience required.

THE PERSON

You will work with a team of Software Engineers to enable DL models, libraries, and applications for Instinct GPUs in both on‑premise and cloud environments. Candidates should be strong in Python and/or C++. Candidates should also have experience analysing and optimising the performance of AI software, understand hardware bottlenecks and harness performance to hit close to roofline. You must be self‑motivated and possess the ability to work well within a team environment.

KEY QUALIFICATIONS
  • Strong programming skills in C++ and Python
  • Strong development experience with at least one major DL framework such as PyTorch or TensorFlow in inference, fine‑tuning and/or training
  • M.S. with years of related experience or Ph.D. with years of related experience in Computer Science, Computer Engineering or a related equivalent
  • Experience developing software and system‑level performance optimisations with a solid architecture understanding in GPUs – a plus
  • Experience with open‑source software development, including collaboration with community maintainers and submitting contributions – a plus
  • Publications in reputed peer‑reviewed ML conferences/journals – a plus
  • Excellent analytical and problem‑solving skills for root‑causing and addressing performance issues
  • Ability to work independently and as part of a team
  • Willingness to learn skills, tools, and methods to advance the quality, consistency, and timeliness of AMD software products
PREFERRED EXPERIENCE
  • Expertise in profiling tools across the AI SW stack (TorchProfiler, RocM profiler, VTune, Nsight)
  • Experience implementing and optimising parallel methods on GPU accelerators (NCCL/RCCL, OpenMP, MPI)
  • Performance analysis skills for both CPU and GPU
  • Experience with Singularity, Docker, and/or Kubernetes
  • Experience providing clear and timely communication related to status and other key aspects of the project to leadership team

This role is not eligible for visa sponsorship.

BENEFITS

Benefits offered are described: AMD benefits at a glance.

EQUAL OPPORTUNITY STATEMENT

AMD and its subsidiaries are equal‑opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third‑party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.

APPLICATION PROCESS

We do not accept unsolicited resumes from headhunters, recruitment agencies, or fee‑based recruitment services. AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's “Responsible AI Policy” is available here. This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Fellow, AI Performance Software Engineer
Fellow, AI Performance Software Engineer

Advanced Micro Devices • Santa Clara (CA)

On-site
USD 100,000 - 140,000
AMD benefits at a glance
Fellow, AI Performance Software Engineer
Fellow, AI Performance Software Engineer

AMD • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Comprehensive benefits package
Inclusive culture
Career advancement opportunities
Staff Software Development Engineer: GPU, Computer Vision, AI/ML Ops
Staff Software Development Engineer: GPU, Computer Vision, AI/ML Ops

AMD • Santa Clara (CA)

On-site
USD 180,000 - 260,000
Benefits at a glance
Fellow, AI Workload Optimization
Fellow, AI Workload Optimization

Advanced Micro Devices • Bellevue (WA)

On-site
USD 180,000 - 220,000
Collaborative culture
Comprehensive benefits package
Software Engineer- GPU/AI/ML
Software Engineer- GPU/AI/ML

AMD • Santa Clara (CA)

On-site
USD 170,000 - 250,000
AMD Benefits
Staff Software Development Engineer: GPU, Computer Vision, AI/ML Ops
Staff Software Development Engineer: GPU, Computer Vision, AI/ML Ops

Advanced Micro Devices • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Fellow, AI Workload Optimization
Fellow, AI Workload Optimization

Advanced Micro Devices, Inc. • Bellevue (WA)

On-site
USD 180,000 - 230,000
Comprehensive benefits package
Inclusive workplace culture
Principal Software Development Eng. - AI Performance
Principal Software Development Eng. - AI Performance

Advanced Micro Devices • San Jose (CA)

On-site
USD 190,000 - 230,000
Fellow Software Engineer - AI Performance & Reliability
Fellow Software Engineer - AI Performance & Reliability

Advanced Micro Devices • San Jose (CA)

On-site
USD 180,000 - 240,000
Opensource Al workload Software Engineer
Opensource Al workload Software Engineer

Socket.dev • San Jose (CA)

On-site
USD 180,000 - 260,000