High Performance and Scientific Computing Fellow

Advanced Micro Devices

Tennessee

On-site

USD 180,000 - 280,000

Full time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Advanced Micro Devices in Austin, TX seeks a High Performance and Scientific Computing Fellow to lead HPC GPU library strategy within the ROCm platform. This IC role emphasizes deep technical vision, algorithms, and performance optimization for next-generation GPUs.

You will mentor engineers, drive collaborations with partners and customers, and shape benchmarks and external engagements to push the boundaries of HPC software across AMD's GPU and CPU stack.

Qualifications

  • In-depth understanding of mathematical algorithms for dense and sparse linear algebra.
  • Experience with direct and iterative solvers.
  • Experience designing, implementing, debugging, and optimizing parallel algorithms on class-leading supercomputers.
  • Knowledge of important HPC algorithms and libraries.
  • Strong background developing applications and libraries in C++, C and Fortran.
  • Familiarity with GPU software development and optimization using HIP, CUDA, or OpenCL and distributed programming with MPI and/or SHMEM.
  • Understanding of CPU and GPU architectures and low-level optimization techniques including assembly programming and vectorization.
  • In-depth knowledge of best-practices in software development, including testing, profiling, debugging, documentation, version control, issue tracking, and planning.

Responsibilities

  • Define technical vision and drive software strategy across AMD ROCm HPC libraries.
  • Lead performance analysis, tuning, and algorithmic innovation across HPC libraries.
  • Provide technical mentorship to senior engineers and influence best practices across the organization
  • Communicate complex technical findings and recommendations to senior leadership and stakeholders.
  • Represent AMD in external technical forums, benchmarks, and customer engagements
  • Work with key technical experts across AMD and with our partners and customers to improve ROCm HPC and AI applications, libraries, and tools.

Skills

Dense linear algebra
High performance computing
C++
Fortran
HIP/CUDA/OpenCL
MPI/SHMEM
GPU architecture
Software optimization
Technical leadership

Education

B.Sc./B.Eng. in CS/EE/Applied Math
M.Sc./M.Eng./Ph.D. preferred

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMDis shapingthefuture.

Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger— technologythat moves the world forward.Join us and, together, we’ll advance your career.

THE ROLE:

The AI GPU Software (AGS) team is looking for a High Performance and Scientific Computing Fellow. This is a technical leadership position as an Individual Contributor (IC) driving innovation across multiple teams developing High Performance Computing (HPC) GPU libraries as part of AMD ROCm™ Software platform (https://github.com/ROCm). You will focus primarily on deep technical leadership rather than organizational management.

As a Fellow, you will be a primary contributor towards defining software strategy and driving the technical vision for HPC GPU libraries across multiple generations of AMD GPUs. This role will require deep expertise in mathematical algorithms, floating point precision, GPU performance analysis, distributed systems, software architecture and engineering. The ideal candidate is highly hands-on and embraces agentic AI workflows and is expected to play a key role in industry-standard benchmarks and external technical engagements.

THE PERSON:

You are accustomed to working in a dynamic, geographically distributed agile team, where partnership and collaboration are paramount. Youpossessexcellent written and verbal communication skills, strong attention to detail, and the ability to express your work in a clear, cohesive fashion. You are a recognized technical leader with contributions across the HPC software stack and applications, both at the node and distributed level, and understand how to architect optimal software solutions for HPC customers while balancing often conflicting priorities. You are comfortable operating across layers—from kernels and runtimes to libraries and distributed strategies—and have a track record of driving impactful optimizations and influencing technical direction.

KEY RESPONSIBILITIES:
  • Define technical vision and drive software strategy across AMD ROCm HPC libraries.
  • Lead performance analysis, tuning, and algorithmic innovation across HPC libraries.
  • Provide technical mentorship to senior engineers and influence best practices across the organization
  • Communicate complex technical findings and recommendations to senior leadership and stakeholders.
  • Represent AMD in external technical forums, benchmarks, and customer engagements
  • Work with key technical experts across AMD and with our partners and customers to improve ROCm HPC and AI applications, libraries, and tools.
PREFERRED EXPERIENCE:
  • In depth understanding of mathematical algorithms for dense and sparse linear algebra
  • Experience with direct and iterative solvers
  • Experience designing, implementing, debugging, and optimizing parallel algorithms on class-leading supercomputers
  • Knowledge of important HPC algorithms and libraries
  • Strong background developing applications and libraries in C++, C and Fortran
  • Familiarity with GPU software development and optimization using HIP, CUDA, or OpenCL and distributed programming with MPI and/or SHMEM
  • Understanding of CPU and GPU architectures and low-level optimization techniques including assembly programming and vectorization
  • In-depth knowledge of best-practices in software development, including testing, profiling, debugging, documentation, version control, issue tracking, and planning
PREFERRED ACADEMIC CREDENTIALS:
  • B.Sc. or B.Eng. degree in Computer Science, Software Engineering, Electrical Engineering, Applied Mathematics, or equivalent
  • Advanced degrees, such as M.Sc., M.Eng., Ph.D. are preferred.
LOCATION:

Austin, TX

This role is not eligible for visa sponsorship.

Benefits offered are described: AMD benefits at a glance.

AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s "Responsible AI Policy" is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

High Performance and Scientific Computing Fellow
High Performance and Scientific Computing Fellow

Socket.dev • Tennessee

Hybrid
USD 200,000 - 320,000
Benefits at a glance
High Performance and Scientific Computing Fellow
High Performance and Scientific Computing Fellow

Advanced Micro Devices, Inc. • Nashville (TN)

On-site
USD 180,000 - 260,000
Principal Software Developer - GPU AI/HPC kernels
Principal Software Developer - GPU AI/HPC kernels

Advanced Micro Devices • Austin (TX)

On-site
USD 130,000 - 160,000
Principal Data Center GPU Performance Architect
Principal Data Center GPU Performance Architect

Advanced Micro Devices • Austin (TX)

On-site
USD 180,000 - 250,000
Principal Datacenter GPU Performance Architect
Principal Datacenter GPU Performance Architect

Advanced Micro Devices • Austin (TX)

On-site
USD 130,000 - 200,000
GPU Runtime and System Software Engineer (Fellow)
GPU Runtime and System Software Engineer (Fellow)

AMD • San Jose (CA)

On-site
USD 230,000 - 420,000
AMD benefits
GPU Kernel Development Engineer (Fellow)
GPU Kernel Development Engineer (Fellow)

Advanced Micro Devices • Austin (TX), Northern (KY)

On-site
USD 210,000 - 300,000
GPU Kernel Development Engineer (Fellow)
GPU Kernel Development Engineer (Fellow)

AMD • Austin (TX)

On-site
USD 190,000 - 270,000
Principal Software Developer - GPU AI/HPC kernels
Principal Software Developer - GPU AI/HPC kernels

AMD • Austin (TX)

On-site
USD 130,000 - 160,000
GPU Performance Architect
GPU Performance Architect

Socket.dev • Santa Clara (CA)

On-site
USD 180,000 - 260,000