Senior GPU Systems Engineer: AI & HPC Platforms

Support Revolution

San Jose (CA)

On-site

USD 137,000 - 156,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Supermicro is seeking an experienced Senior Systems Engineer / GPU Platforms to support bring‑up, qualification, and customer deployment of multi‑GPU server platforms for AI, HPC, and enterprise workloads.

The candidate will work across the product lifecycle, partner with Architecture, Systems, Software, and Validation teams, and provide technical leadership, documentation, and training to peers and customers.

Qualifications

  • Bachelor's degree in Engineering/CS/IT or equivalent practical experience.
  • 5–15 years of relevant industry experience in systems engineering, server engineering, validation or AI infrastructure.
  • Strong knowledge of enterprise server hardware and system architecture.
  • Hands‑on experience with Linux server environments.
  • Experience installing, configuring, validating, and troubleshooting server hardware and software.
  • Strong system‑level troubleshooting and root‑cause analysis skills.
  • Working knowledge of PCIe architectures and high‑performance I/O.
  • Experience with GPU computing, accelerators, or HPC technologies.
  • Ability to manage complex technical assignments and drive issues to resolution.
  • Strong written and verbal communication skills.
  • Ability to work with cross‑functional and distributed teams.
  • Comfortable participating in customer‑facing technical discussions.

Responsibilities

  • Assist bring‑up, configuration, integration, validation, and troubleshooting of advanced GPU server platforms.
  • Execute GPU platform qualification activities including NVQUAL or equivalent.
  • Install, configure, troubleshoot Linux, GPU drivers, CUDA, firmware, libraries and related software.
  • Diagnose complex system issues using logs, telemetry, diagnostics and vendor tools.
  • Support multi‑GPU server platforms through qualification, launch and post‑release activities.
  • Engage in customer‑facing POC/EVAL engagements with system preparation and debugging.
  • Collaborate with Architecture, Systems, Software, Validation, Product Management and partners.
  • Develop technical documentation, troubleshooting guides and best practices.
  • Deliver technical presentations and training sessions; mentor others as needed.
  • Maintain ownership and accountability for complex technical issues.
  • Learn new server/GPU/software technologies quickly and effectively.
  • Support cross‑functional collaboration and knowledge sharing.

Skills

Linux server environments
Server hardware knowledge
System‑level troubleshooting
Cross‑functional collaboration
Written and verbal communication

Education

Bachelor's degree in Engineering/CS/IT

Tools

NVIDIA GPU software
CUDA
NVQUAL
Docker
Kubernetes
PCIe architectures knowledge

Job description

Supermicro is seeking an experienced Senior Systems Engineer / GPU Platforms to support bring‑up, qualification, and customer deployment of multi‑GPU server platforms for AI, HPC, and enterprise workloads.

The candidate will work across the product lifecycle, partner with Architecture, Systems, Software, and Validation teams, and provide technical leadership, documentation, and training to peers and customers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior GPU Platform Engineer
Senior GPU Platform Engineer

Supermicro • Wayne (CA)

On-site
USD 137,000 - 156,000
Sr. System Engineer/GPU Platforms
Sr. System Engineer/GPU Platforms

Supermicro • Wayne (CA)

On-site
USD 137,000 - 156,000
Sr. System Engineer/GPU Platforms
Sr. System Engineer/GPU Platforms

Support Revolution • San Jose (CA)

On-site
USD 137,000 - 156,000
GPU Server Systems Engineer — Data Center & Automation
GPU Server Systems Engineer — Data Center & Automation

Supermicro • San Jose (CA)

On-site
USD 90,000 - 110,000
Data Center Solutions Engineer: GPU & AI Infra
Data Center Solutions Engineer: GPU & AI Infra

Supermicro • San Jose (CA)

On-site
USD 165,000 - 200,000
Comprehensive benefits
Bonus and equity programs
Senior HPC & AI Cluster Engineer
Senior HPC & AI Cluster Engineer

Support Revolution • San Jose (CA), Northern (KY)

Hybrid
USD 137,000 - 156,000
Staff Data Center Solutions Engineer - GPU Cluster Expert
Staff Data Center Solutions Engineer - GPU Cluster Expert

Super Micro Computer Spain, S.L. • San Jose (CA)

On-site
USD 165,000 - 200,000
System Engineer, GPU Server
System Engineer, GPU Server

Super Micro Computer Spain, S.L. • San Jose (CA)

On-site
USD 90,000 - 110,000
GPU Server Systems Engineer – Data Center & Benchmarks
GPU Server Systems Engineer – Data Center & Benchmarks

Super Micro Computer Spain, S.L. • San Jose (CA)

On-site
USD 90,000 - 110,000
System Engineer, GPU Server
System Engineer, GPU Server

Supermicro • San Jose (CA)

On-site
USD 90,000 - 110,000