Senior HPC & AI Cluster Engineer

Support Revolution

San Jose, Northern (CA, KY)

Hybrid

USD 137,000 - 156,000

Full time

9 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Supermicro is a global leader in server technologies seeking a Senior System Engineer to design, implement and deploy rack-scale solutions for data centers and enterprise customers. You will lead deployment, testing and optimization of compute, storage and networking, author test procedures, and provide on-site deployment services.

The role emphasizes HPC/AI benchmarks, Python scripting, and collaboration with cross-functional teams.

Qualifications

  • BS/MS in Electrical or Computer Engineering or related field.
  • 8+ years of server/hardware testing, deployment, and troubleshooting.
  • 8+ years in DevOps or cloud environments (Docker/Kubernetes).
  • Experience with AI/ML frameworks (PyTorch, TensorFlow).
  • Strong Python and shell scripting abilities.

Responsibilities

  • Deploy Rack/Cluster infrastructure and perform comprehensive system testing.
  • Lead design validation, HPC/AI benchmarks, and OS/network tuning.
  • Provide on-site deployment services and collaborate with engineering teams.
  • Write test procedures, reports, and troubleshooting documents.
  • Develop automation tools for cluster deployment and testing.

Skills

Server hardware configuration
DevOps / Cloud
Python programming
Teamwork

Education

BS/MS in Electrical/Computer Engineering

Tools

Docker
Kubernetes
OpenStack
Azure
AWS

Job description

Supermicro is a global leader in server technologies seeking a Senior System Engineer to design, implement and deploy rack-scale solutions for data centers and enterprise customers. You will lead deployment, testing and optimization of compute, storage and networking, author test procedures, and provide on-site deployment services.

The role emphasizes HPC/AI benchmarks, Python scripting, and collaboration with cross-functional teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior System Engineer, HPC/AI & Cloud Clusters
Senior System Engineer, HPC/AI & Cloud Clusters

Supermicro • San Jose (CA)

On-site
USD 137,000 - 156,000
Sr. System Engineer
Sr. System Engineer

Supermicro • San Jose (CA)

On-site
USD 137,000 - 156,000
Sr. System Engineer
Sr. System Engineer

Support Revolution • San Jose (CA), Northern (KY)

On-site
USD 137,000 - 156,000
Senior GPU Platform Engineer for AI/HPC
Senior GPU Platform Engineer for AI/HPC

Super Micro Computer Spain, S.L. • San Jose (CA)

On-site
USD 137,000 - 156,000
Data Center Software Engineer – AI & Automation
Data Center Software Engineer – AI & Automation

Support Revolution • San Jose (CA)

On-site
USD 100,000 - 115,000
Senior Data Center Engineer — Systems & Infrastructure
Senior Data Center Engineer — Systems & Infrastructure

Support Revolution • San Jose (CA)

On-site
USD 105,000 - 130,000
System Engineer - Data Center Benchmarks & Server Tech
System Engineer - Data Center Benchmarks & Server Tech

Support Revolution • San Jose (CA)

On-site
USD 85,000 - 100,000
Staff Data Center Solutions Engineer - GPU AI Deployments
Staff Data Center Solutions Engineer - GPU AI Deployments

Supermicro • San Jose (CA)

On-site
USD 165,000 - 200,000
Senior Systems Engineer - Networking & Data Center
Senior Systems Engineer - Networking & Data Center

Supermicro • San Jose (CA)

On-site
USD 137,000 - 156,000
Senior Data Center & HPC Solutions Sales Leader
Senior Data Center & HPC Solutions Sales Leader

Supermicro • San Jose (CA)

On-site
USD 110,000 - 178,000