HPC & AI Network Engineer – GPU & Cloud

Support Revolution

San Jose (CA)

On-site

USD 120,000 - 140,000

Full time

8 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Supermicro is seeking a Network Engineer to support rollouts and ongoing maintenance of business-critical applications and services. You will collaborate with senior engineers to resolve escalated issues and implement complex projects across the data center ecosystem.

You will perform system-level tests on latest GPUs and processors, tune BIOS and OS/network settings, and help document procedures. A background in DL/ML and Linux networking is highly valued for this role.

Qualifications

  • For applicants with a degree, at least 1+ years of relevant work in DL/ML and related research.
  • Experience with Linux networking debugging or testing is preferred.
  • Familiarity with data center, enterprise, or telecom networking technologies.
  • Hands-on experience with workloads and cluster management (Slurm).
  • Experience with MLPerf benchmarks, LLM, RCCL/NCCL or similar runtimes.
  • Proficient in Windows and Linux shell scripting; strong communication skills.

Responsibilities

  • Execute system-level rack tests on GPUs, CPUs, and accelerators to verify functionality and reliability.
  • Address HPC/AI customer issues and develop robust processes for HPC/AI solutions.
  • Contribute to proof-of-concept designs and optimized benchmarks for HPC/AI workloads.
  • Deliver on-site deployment services and provide level 1–2 support.
  • Create documentation, notes, blogs, and diagrams to share technical knowledge.
  • Collaborate with Product and Engineering teams to feed customer feedback into product improvements.
  • Engage in HPC roadmap planning and plan software/hardware upgrades.
  • Document test plans, results, and automate testing where possible.

Skills

Deep Learning
Machine Learning
Linux Networking
Slurm
MLPerf
NCCL
Shell scripting
Teamwork

Education

BS/MS in Electrical Engineering, Computer Engineering or Computer Science

Tools

Docker
Kubernetes

Job description

Supermicro is seeking a Network Engineer to support rollouts and ongoing maintenance of business-critical applications and services. You will collaborate with senior engineers to resolve escalated issues and implement complex projects across the data center ecosystem.

You will perform system-level tests on latest GPUs and processors, tune BIOS and OS/network settings, and help document procedures. A background in DL/ML and Linux networking is highly valued for this role.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Network & HPC Systems Engineer
Senior Network & HPC Systems Engineer

Super Micro Computer, Inc. • San Jose (CA)

On-site
USD 120,000 - 140,000
HPC Network Systems Engineer for AI & Data Center
HPC Network Systems Engineer for AI & Data Center

Supermicro • Wayne (CA)

On-site
USD 120,000 - 140,000
HPC/AI Network Engineer — Equity Eligible
HPC/AI Network Engineer — Equity Eligible

Supermicro • San Jose (CA)

On-site
USD 120,000 - 140,000
Data Center Network Engineer — AI/GPU Lab Connectivity
Data Center Network Engineer — AI/GPU Lab Connectivity

Support Revolution • San Jose (CA)

Hybrid
USD 115,000 - 130,000
Senior System Engineer – HPC/AI & Cloud Deployments
Senior System Engineer – HPC/AI & Cloud Deployments

Support Revolution • San Jose (CA), Northern (KY)

Hybrid
USD 137,000 - 156,000
Senior Data Center Solutions Engineer - GPU & AI Infra
Senior Data Center Solutions Engineer - GPU & AI Infra

Support Revolution • San Jose (CA)

On-site
USD 165,000 - 200,000
Senior GPU Platform Engineer
Senior GPU Platform Engineer

Supermicro • Wayne (CA)

On-site
USD 137,000 - 156,000
Network Systems Engineer
Network Systems Engineer

Supermicro • Wayne (CA)

On-site
USD 120,000 - 140,000
Senior System Engineer, HPC/AI & Cloud Clusters
Senior System Engineer, HPC/AI & Cloud Clusters

Supermicro • San Jose (CA)

On-site
USD 137,000 - 156,000
Network Systems Engineer
Network Systems Engineer

Supermicro • San Jose (CA)

On-site
USD 120,000 - 140,000