Cloud Storage Hardware Reliability Engineer

Amazon

Seattle (WA)

On-site

USD 136,000 - 184,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Amazon Storage Server Hardware Engineering seeks a Cloud Hardware Development Engineer to own fleet reliability and sustaining engineering for deployed storage platforms.

You will drive failure analysis, component lifecycle management, and feed-forward loop between field performance and next-generation design, partnering with ODMs, firmware, and operations to maintain end-to-end platform quality in data center environments.

Qualifications

  • Bachelor's degree in Electrical Engineering, Computer Engineering, or equivalent.
  • 2+ years hardware design, development and validation experience for server or compute platforms.
  • Experience in one or more server technologies: thermal/mechanical design, power delivery, high-speed signal integrity, or accelerator subsystems.
  • Experience developing functional specifications, design verification plans, and validation test procedures.

Responsibilities

  • Own fleet quality metrics post-launch: platform-level annualized failure rates, unsellable server rates, and component-level failure modes.
  • Drive root cause analysis on high-volume failure patterns — particularly connector and cable-driven failures, storage controller faults, and power-loss protection issues.
  • Correlate field failure data with manufacturing lot, supplier, and test escape information to identify systemic trends.
  • Engage with component suppliers on corrective actions.
  • Manage platform-specific BOM variants across multiple deployed generations simultaneously.
  • Translate fleet failure patterns into concrete requirements for next-generation platforms.
  • Define test content additions based on failure modes observed in production.
  • Contribute failure-mode-driven additions to manufacturing test coverage and data center vetting automation.
  • Define and execute validation strategies from PCBA bring-up through server and rack integration.
  • Own hardware debug during EVT/DVT/PVT builds, correlating failures across PCIe, power rails, NVMe links, and storage controller subsystems.
  • Triage hardware issues at ODM facilities and datacenters, conduct root cause analysis, and implement corrective actions.
  • Work with firmware, software, and operations teams to ensure diagnostic tooling and fleet dashboards provide visibility into fleet health.
  • May require occasional (<10%) regional and international travel to Design and Manufacturing Partner sites.

Skills

Hardware design
Validation experience
Server technologies
Power delivery
Thermal design
NVMe/SSD subsystems
PCIe topology
Root cause analysis
Failure analysis

Education

Bachelor's degree in Electrical Engineering, Computer Engineering, or equivalent
Master’s degree (preferred)

Tools

ODM coordination
Design verification plans
Validation test procedures

Job description

Amazon Storage Server Hardware Engineering seeks a Cloud Hardware Development Engineer to own fleet reliability and sustaining engineering for deployed storage platforms.

You will drive failure analysis, component lifecycle management, and feed-forward loop between field performance and next-generation design, partnering with ODMs, firmware, and operations to maintain end-to-end platform quality in data center environments.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Cloud Storage Hardware Reliability Engineer
Cloud Storage Hardware Reliability Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 157,000 - 213,000
RSUs
Health insurance
401(k) matching
Cloud Storage Hardware Engineer – Fleet Reliability
Cloud Storage Hardware Engineer – Fleet Reliability

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 136,000 - 184,000
Storage Server Hardware Engineer - NPI & Fleet Reliability
Storage Server Hardware Engineer - NPI & Fleet Reliability

Amazon Web Services (AWS) • Denver (CO)

On-site
USD 159,000 - 215,000
Cloud Storage Server Hardware Engineer
Cloud Storage Server Hardware Engineer

Amazon Web Services (AWS) • Denver (CO)

On-site
USD 159,000 - 215,000
Health insurance
401(k) matching
Parental leave
Senior Hardware Design Engineer - SSD & Storage for Cloud
Senior Hardware Design Engineer - SSD & Storage for Cloud

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 157,000 - 213,000
Lead Hardware Design Engineer, SSD & Cloud Systems
Lead Hardware Design Engineer, SSD & Cloud Systems

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 136,000 - 184,000
Health insurance
401(k) matching
Paid time off
+1
Senior HDD Hardware Engineer for Cloud Storage
Senior HDD Hardware Engineer for Cloud Storage

Amazon Web Services (AWS) • Denver (CO)

On-site
USD 159,000 - 215,000
Comprehensive benefits
RSU eligibility
Datacenter Reliability Engineer, Hardware & Quality
Datacenter Reliability Engineer, Hardware & Quality

Amazon • Herndon (VA)

On-site
USD 117,000 - 160,000
Lead Hardware Design Engineer — SSD Storage (Cloud)
Lead Hardware Design Engineer — SSD Storage (Cloud)

Amazon • Cupertino (CA)

On-site
USD 157,000 - 213,000
Health insurance
Stock-based compensation (RSUs)
401(k) matching
Cloud Server Hardware Engineer
Cloud Server Hardware Engineer

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 70,000 - 116,000