RMA Failure Analysis Engineer: GPU Server Systems

CoreFleet Solutions

San Jose (CA)

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Growth opportunities
Direct leadership interaction
Skill development
Collaborative culture
Impactful work

Job summary

CoreFleet Solutions is seeking an experienced RMA Failure Analysis engineer for GPU servers and enterprise server platforms. You will diagnose, troubleshoot, and perform root-cause analysis on customer-returned GPU servers, server motherboards, GPU boards, and related hardware subsystems.

The role requires strong server architecture knowledge, component-level debugging skills, and the ability to safely handle high-value hardware while conducting failure analysis.

Qualifications

  • 4 years of experience in server hardware design, validation, testing, failure analysis, or system engineering.
  • Strong understanding of GPU server architectures.
  • Experience reading electrical schematics, block diagrams, and PCB layouts.
  • Hands-on experience with lab equipment and troubleshooting across BIOS, BMC, PCIe, memory, storage, and networking.

Responsibilities

  • Perform failure analysis on customer-returned GPU servers and related hardware.
  • Conduct system-, board-, and component-level troubleshooting to identify root causes.
  • Execute functional testing and diagnostics using standard lab equipment.
  • Read schematics, block diagrams, layouts, and manufacturing documentation.
  • Document failure analysis findings and communicate with engineering, quality, and customer support teams.
  • Follow proper ESD and hardware handling procedures.

Skills

Server hardware design
Failure analysis
System engineering
Root cause analysis
Linux troubleshooting

Tools

Oscilloscopes
Digital multimeters
Power analyzers
Logic analyzers
Protocol analyzers

Job description

CoreFleet Solutions is seeking an experienced RMA Failure Analysis engineer for GPU servers and enterprise server platforms. You will diagnose, troubleshoot, and perform root-cause analysis on customer-returned GPU servers, server motherboards, GPU boards, and related hardware subsystems.

The role requires strong server architecture knowledge, component-level debugging skills, and the ability to safely handle high-value hardware while conducting failure analysis.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RMA Failure Analysis Engineer
RMA Failure Analysis Engineer

CoreFleet Solutions • San Jose (CA)

On-site
USD 120,000 - 180,000
Growth opportunities
Direct leadership interaction
Skill development
+2
GPU Server Failure Analysis Engineer – Data Center Debug
GPU Server Failure Analysis Engineer – Data Center Debug

Advanced Micro Devices, Inc. • Austin (TX)

On-site
USD 120,000 - 150,000
Senior GPU Server Failure Analysis Engineer
Senior GPU Server Failure Analysis Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 95,000 - 130,000
Health insurance
Paid time off
Professional development
System Failure Analysis Engineer (GPU Servers / Data Center)
System Failure Analysis Engineer (GPU Servers / Data Center)

AMD • Austin (TX)

On-site
USD 100,000 - 130,000
RMA Failure Analysis Lead & On-Site Engineer
RMA Failure Analysis Lead & On-Site Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 85,000 - 130,000
AMD benefits
RMA Hardware Troubleshooter - GPU/CPU & Data Center
RMA Hardware Troubleshooter - GPU/CPU & Data Center

Vultr • United States

Remote
USD 75,000 - 90,000
GPU ASIC & PCBA Failure Analysis Engineering Lead
GPU ASIC & PCBA Failure Analysis Engineering Lead

Advanced Micro Devices, Inc. • Secaucus (NJ)

On-site
USD 130,000 - 160,000
Competitive benefits
GPU & PCBA Failure Analysis Engineering Manager
GPU & PCBA Failure Analysis Engineering Manager

Advanced Micro Devices • Secaucus (NJ)

On-site
USD 120,000 - 150,000
Failure Analysis Engineer
Failure Analysis Engineer

Ultimate Staffing Services • Santa Clara (CA)

On-site
USD 48,000 - 62,000
IT Infrastructure Engineer – RMA & Hardware Diagnostics
IT Infrastructure Engineer – RMA & Hardware Diagnostics

Nebius B.V. • Kansas City (MO)

On-site
USD 90,000 - 130,000