AI/ML Server Hardware Architect in High-Performance Systems

Amazon Web Services (AWS)

Cupertino (CA)

On-site

USD 183,000 - 248,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401(k) matching
RSU equity

Job summary

Amazon Web Services (AWS) is seeking a Hardware Design/System Engineer to shape next‑gen server platforms for AI and HPC at scale in Cupertino. You will design, build, and operate high‑performance infrastructure that powers AI training and inference workloads, collaborating across hardware, software, and supply chain teams to deliver reliable systems.

You’ll own diagnostics, testing, and production readiness, working with ODM/JDM partners to bring concepts from concept to production, while

Qualifications

  • Experience in developing functional specifications, design verification plans and functional test procedures.
  • Experience in server technologies such as, thermal, mechanical, power, and signal integrity.
  • Bachelor's degree or above in electrical engineering, computer engineering, or equivalent.
  • 5+ years of Design/Innovation, research & development, manufacturing, process, industrial engineering, or related experience.
  • 5+ years of process development experience.
  • Experience in English-language communication skills, both written and verbal.
  • In depth expertise in server technologies such as Thermal / Mechanical design, high speed bus design and signal integrity, failure analysis, server components (e.g. CPU, GPU, SSDs, memory), BIOS, BMC, and networking

Responsibilities

  • Lead technical solutions for complex high performance server and/or accelerator server and rack system architectural challenges.
  • Own end-to-end system reliability, proactively identifying and resolving deficiencies before customer impact.
  • Design and implement solutions to address system-level issues at large scale.
  • Decompose complex server system problems (testability, reliability, diagnostics) into deliverable tasks and features.
  • Apply expertise across hardware, software, system design, x86 architecture, processes, and operations.
  • Collaborate with hardware, software, manufacturing, supply chain and product management teams.
  • Develop and implement diagnostic tools and monitoring solutions for production systems.
  • Debug complex system failures in time sensitive settings.

Skills

Server technologies
Thermal / Mechanical design
High speed bus design
BIOS / BMC / Networking
English communication
Linux
PowerShell
Python

Education

Bachelor's degree
Master's degree

Tools

Linux
PowerShell
Python

Job description

Amazon Web Services (AWS) is seeking a Hardware Design/System Engineer to shape next‑gen server platforms for AI and HPC at scale in Cupertino. You will design, build, and operate high‑performance infrastructure that powers AI training and inference workloads, collaborating across hardware, software, and supply chain teams to deliver reliable systems.

You’ll own diagnostics, testing, and production readiness, working with ODM/JDM partners to bring concepts from concept to production, while

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI/ML Server Hardware Architect
Senior AI/ML Server Hardware Architect

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 159,200 - 215,300
Cloud AI Hardware Engineer — Scale High‑Perf Servers
Cloud AI Hardware Engineer — Scale High‑Perf Servers

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 183,000 - 248,000
AI/ML GPU Server Hardware Engineer
AI/ML GPU Server Hardware Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 126,000 - 185,000
Health insurance
401(k) matching
Paid time off
+1
Senior AI/ML HPC Systems Engineer - Accelerator Servers
Senior AI/ML HPC Systems Engineer - Accelerator Servers

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 173,000 - 236,000
Sign-on bonuses and RSUs
Generative AI Cloud Hardware Engineer
Generative AI Cloud Hardware Engineer

Amazon • Cupertino (CA)

On-site
USD 157,000 - 213,000
Systems Architect: AWS GenAI & ML Server
Systems Architect: AWS GenAI & ML Server

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 148,000 - 202,000
Sign-on payments
Restricted stock units (RSUs)
Health, dental, vision
+3
Cloud Hardware Engineer — AI/ML Server Fleet
Cloud Hardware Engineer — AI/ML Server Fleet

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 159,000 - 215,000
Sr Hardware Development Engineer, High Performance AI & ML Servers
Sr Hardware Development Engineer, High Performance AI & ML Servers

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 159,200 - 215,300
Cloud AI Hardware Engineer – Accelerated Server Systems
Cloud AI Hardware Engineer – Accelerated Server Systems

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 136,000 - 184,000
Cloud Hardware Engineer for Accelerated AI Servers
Cloud Hardware Engineer for Accelerated AI Servers

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 110,000 - 160,000
Health insurance
401(k) matching
Paid time off