GPU Cluster Performance Engineer - RDMA & Scaling

AMD

Austin (TX)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

AMD in Austin, TX, is seeking a talented GPU Cluster Network Performance Attainment Engineer to optimize GPU clusters for peak performance. The ideal candidate will evaluate scalability, perform benchmarking, and implement performance tuning strategies while collaborating with hardware and software teams.

This position requires a deep understanding of GPU architectures, RDMA networks, and performance optimization methodologies. A degree in electrical or computer engineering is preferred. Benefits include comprehensive AMD perks, and this role does not offer visa sponsorship.

Qualifications

  • Proven experience in optimizing the performance of GPU clusters.
  • Strong understanding of GPU architectures and parallel computing concepts.
  • Experience with network protocols used in GPU clusters.

Responsibilities

  • Evaluate scalability of GPU clusters under various workloads.
  • Utilize profiling tools to analyze performance bottlenecks.
  • Implement optimization strategies for performance tuning.
  • Collaborate with cross-functional teams to enhance GPU cluster performance.
  • Develop benchmarking strategies to assess performance.

Skills

Optimization of GPU clusters
Understanding of RDMA network drivers
Scripting languages (Python, Bash)
System-level performance analysis tools
Excellent communication and collaboration skills
Linux kernel networking expertise
Analytical mindset and problem-solving skills
Machine learning and/or HPC system design

Education

Bachelor’s or Master’s degree in electrical or computer engineering

Job description

AMD in Austin, TX, is seeking a talented GPU Cluster Network Performance Attainment Engineer to optimize GPU clusters for peak performance. The ideal candidate will evaluate scalability, perform benchmarking, and implement performance tuning strategies while collaborating with hardware and software teams.

This position requires a deep understanding of GPU architectures, RDMA networks, and performance optimization methodologies. A degree in electrical or computer engineering is preferred. Benefits include comprehensive AMD perks, and this role does not offer visa sponsorship.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI GPU Cluster Performance Engineer
Senior AI GPU Cluster Performance Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 100,000 - 130,000
Senior AI GPU Cluster Performance Engineer
Senior AI GPU Cluster Performance Engineer

Socket.dev • Austin (TX)

Hybrid
USD 120,000 - 160,000
AMD benefits at a glance
Senior AI Cluster Hardware Engineer
Senior AI Cluster Hardware Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 100,000 - 130,000
Senior AI Cluster Hardware Engineer
Senior AI Cluster Hardware Engineer

AMD • Austin (TX)

On-site
USD 100,000 - 130,000
Senior AI Cluster Hardware Engineer
Senior AI Cluster Hardware Engineer

Socket.dev • Austin (TX)

Hybrid
USD 120,000 - 160,000
AMD benefits at a glance
AI Cluster Validation Engineer
AI Cluster Validation Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 140,000 - 190,000
Senior AI Cluster Validation Engineer
Senior AI Cluster Validation Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 140,000 - 190,000
Data Center GPU Performance Attainment Lead
Data Center GPU Performance Attainment Lead

Advanced Micro Devices • Austin (TX)

Hybrid
USD 90,000 - 120,000
Data Center GPU/CPU Field Engineer
Data Center GPU/CPU Field Engineer

CareerArc • Austin (TX)

On-site
USD 140,000 - 190,000
AI Cluster Program Lead: GPU & Rack Validation
AI Cluster Program Lead: GPU & Rack Validation

Advanced Micro Devices • Austin (TX)

On-site
USD 140,000 - 210,000
AMD Benefits