Senior HPC Infrastructure Engineer (Remote)

Guardant Health

United States

Hybrid

USD 147,000 - 214,000

Full time

7 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Guardant Health is seeking a Staff-level HPC engineer to own and scale our HPC infrastructure, with deep depth in Red Hat Linux, networking, storage, Kubernetes or Slurm. You will collaborate across IT, SQA, DevOps/SRE and MSP to deliver reliable, scalable HPC services for oncology data processing.

The role emphasizes 24/7 on-call coverage, on-site hybrid work in the US, and close partnership with vendors and offshore consultants to drive resilient infrastructure.

Qualifications

  • Bachelor’s degree in Computer Science or related field with 8–12 years of relevant experience; Master’s with 6–8 years; or PhD with 3–5 years.
  • Strong experience in systems/infrastructure engineering, Linux/Unix administration and TCP/IP networking.
  • Hands-on automation with Ansible or equivalent tooling.
  • Experience with high-performance networking tech, including InfiniBand/RDMA and troubleshooting.
  • Experience supporting large-scale storage and HPC/compute environments.
  • Experience in on-prem and cloud-based infrastructure (AWS/GCP/Azure).
  • Experience with software release, operations and infrastructure automation processes, plus strong documentation.

Responsibilities

  • Manage multiple HPC clusters and cluster file systems.
  • Integrate cloud bursting as part of HPC abstraction work.
  • Develop next generation HPC solutions and troubleshoot production stack to code level.
  • Maintain and monitor infrastructure and provide 24/7 on-call rotation.
  • Collaborate across networking, storage, SQA, DevOps/SRE and MSP.
  • Mentor junior engineers on HPC best practices.
  • Work with offsite consultants and vendors to upgrade systems.

Skills

Linux/Unix
TCP/IP networking
Automation (Ansible)
High-performance networking
SRE/DevOps collaboration
Shell scripting
Cross-functional teamwork

Education

Bachelor’s degree in CS or related field
Master’s or PhD in related field (optional)

Tools

Ansible
GPFS
Slurm
Docker
Kubernetes
Warewulf
InfiniBand/ RDMA
Red Hat Linux

Job description

Guardant Health is seeking a Staff-level HPC engineer to own and scale our HPC infrastructure, with deep depth in Red Hat Linux, networking, storage, Kubernetes or Slurm. You will collaborate across IT, SQA, DevOps/SRE and MSP to deliver reliable, scalable HPC services for oncology data processing.

The role emphasizes 24/7 on-call coverage, on-site hybrid work in the US, and close partnership with vendors and offshore consultants to drive resilient infrastructure.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC Infrastructure Engineer – Remote
Senior HPC Infrastructure Engineer – Remote

Guardant Health, Inc. • United States

Hybrid
USD 156,000 - 214,000
Staff HPC Infrastructure Engineer
Staff HPC Infrastructure Engineer

Guardant Health • United States

Hybrid
USD 147,000 - 214,000
Staff HPC Infrastructure Engineer
Staff HPC Infrastructure Engineer

Guardant Health, Inc. • United States

Hybrid
USD 156,000 - 214,000
Senior HPC & Cloud Systems Lead (Remote)
Senior HPC & Cloud Systems Lead (Remote)

RedLine Performance Solutions, LLC • Silver Spring (MD)

On-site
USD 120,000 - 180,000
Paid time off
401k match
Health care benefits
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters

Parallel Works • Chicago (IL)

Hybrid
USD 140,000 - 190,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Senior Linux & HPC Infra Engineer — Hybrid (Remote 2 days)
Senior Linux & HPC Infra Engineer — Hybrid (Remote 2 days)

FinOps Weekly • Center (TX), Northern (KY)

Hybrid
USD 120,000 - 150,000
Senior HPC Infrastructure Engineer: Clusters & Cloud
Senior HPC Infrastructure Engineer: Clusters & Cloud

Jobtailor • California (MO)

On-site
USD 150,000 - 210,000
Senior HPC Linux Systems Engineer - Hybrid/On-Site
Senior HPC Linux Systems Engineer - Hybrid/On-Site

UT-Battelle • Oak Ridge (TN)

On-site
USD 120,000 - 180,000
Senior Security Engineer - Hybrid (Remote/Onsite)
Senior Security Engineer - Hybrid (Remote/Onsite)

Guardant Health • Palo Alto (CA)

Hybrid
USD 130,000 - 180,000
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux

Strategic Business Systems, Inc (SBS) • Chantilly (VA)

Hybrid
USD 120,000 - 180,000
Flexible work arrangements