GPU Server Hardware Engineer for AI/ML

Amazon

Cupertino (CA)

On-site

USD 126,000 - 185,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

RSUs
Health benefits
401(k) matching

Job summary

Amazon Data Services, Inc. is seeking a Hardware Development Engineer to join the AI/ML server development team in Cupertino. You will investigate hardware issues on GPU server platforms, analyze failure data, and collaborate with senior engineers to understand server architecture and fleet operations.

The role offers growth through mentorship and hands-on cross-stack exposure. You'll contribute to design reviews, document findings, and help drive reliability improvements across hardware,

Qualifications

  • Bachelor's degree or above in electrical engineering, computer engineering, or equivalent.
  • Experience communicating technical concepts using clear visuals.
  • Understanding of computer architecture (CPU, memory, buses, I/O).
  • Hands-on experience with Linux environments.
  • Strong organizational skills and cross-functional teamwork.
  • Desire and energy to work in a fast-paced, learning-intensive environment.

Responsibilities

  • Investigate and root cause hardware issues on GPU server platforms with support of an experienced engineer.
  • Analyze failure data and trends to identify systemic problems and propose fixes.
  • Work with senior engineers to understand server architecture, GPU subsystems, and fleet operations.
  • Participate in design reviews for new server platforms, contributing a reliability and serviceability perspective.
  • Collaborate with cross-functional teams to close the loop between field failures and design improvements.
  • Document findings, contribute to runbooks, and share knowledge with the team.

Skills

Python
Java
C/C++
Linux

Education

Bachelor's degree in EE/CE or related field

Job description

Amazon Data Services, Inc. is seeking a Hardware Development Engineer to join the AI/ML server development team in Cupertino. You will investigate hardware issues on GPU server platforms, analyze failure data, and collaborate with senior engineers to understand server architecture and fleet operations.

The role offers growth through mentorship and hands-on cross-stack exposure. You'll contribute to design reviews, document findings, and help drive reliability improvements across hardware,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Server Hardware Engineer for AI Systems
GPU Server Hardware Engineer for AI Systems

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 110,000 - 160,000
Cloud Hardware Engineer: AI/ML/GPU Server Systems
Cloud Hardware Engineer: AI/ML/GPU Server Systems

Amazon • Cupertino (CA)

On-site
USD 157,000 - 213,000
Comprehensive benefits including RSU/"
Generative AI Cloud Hardware Engineer
Generative AI Cloud Hardware Engineer

Amazon • Cupertino (CA)

On-site
USD 157,000 - 213,000
Cloud AI Systems Engineer - Generative AI & ML Servers
Cloud AI Systems Engineer - Generative AI & ML Servers

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 149,000 - 201,000
Health insurance
401(k) matching
Paid time off
AI/ML Server Hardware Architect in High-Performance Systems
AI/ML Server Hardware Architect in High-Performance Systems

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 183,000 - 248,000
Health insurance
401(k) matching
RSU equity
Cloud Hardware Engineer for AI Acceleration Servers
Cloud Hardware Engineer for AI Acceleration Servers

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 126,000 - 185,000
RSUs
Competitive compensation
Cloud AI/ML Hardware Engineer — Storage & Server Platforms
Cloud AI/ML Hardware Engineer — Storage & Server Platforms

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 157,000 - 213,000
Hardware Development Engineer, AI/ML Server Development
Hardware Development Engineer, AI/ML Server Development

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 110,000 - 160,000
Cloud AI Hardware Engineer — Scale High‑Perf Servers
Cloud AI Hardware Engineer — Scale High‑Perf Servers

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 183,000 - 248,000
Cloud Hardware Engineer for AI/ML Accelerator Servers
Cloud Hardware Engineer for AI/ML Accelerator Servers

Amazon • Austin (TX)

On-site
USD 110,000 - 160,000
Health insurance
RSUs
401(k) matching