Senior HPC Infrastructure Engineer

Guardant Health

Palo Alto (CA)

Hybrid

USD 173,000 - 238,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Hybrid work model

Job summary

Guardant Health is seeking a Staff HPC Engineer in Palo Alto to lead and scale high-performance compute environments. You will manage multiple HPC clusters, integrate cloud bursting, and troubleshoot the production stack from shell scripts to code, while collaborating with networking, storage, and MSP teams.

The role requires 8–12 years of systems engineering experience, deep Linux knowledge, and hands-on work with Slurm, Kubernetes, GPFS, and cloud platforms.

Qualifications

  • Bachelor’s degree in Computer Science or a related field with 8–12 years of relevant experience; Master’s degree with 6–8 years; or PhD with 3–5 years.
  • Hands-on experience with Linux/Unix administration and TCP/IP networking.
  • Experience with automation tools such as Ansible or equivalent technologies.
  • Experience with high-performance networking technologies (InfiniBand, RoCE, RDMA).
  • Experience supporting large-scale HPC/compute environments and on-premise and cloud infrastructure.

Responsibilities

  • Manage multiple HPC clusters and cluster file systems.
  • Integrate cloud bursting as part of HPC abstraction work.
  • Troubleshoot production stack to source code level.
  • Maintain and monitor infrastructure and tooling.
  • Support remote locations, including international sites.
  • Mentor junior engineers on HPC best practices.

Skills

Linux administration
HPC
Networking
Slurm
Kubernetes
Storage
Ansible
AWS

Education

Bachelor’s degree in Computer Science
Master’s degree in Computer Science
PhD

Tools

Slurm
GPFS
Warewulf
Ansible
Kubernetes
AWS
GCP

Job description

Guardant Health is seeking a Staff HPC Engineer in Palo Alto to lead and scale high-performance compute environments. You will manage multiple HPC clusters, integrate cloud bursting, and troubleshoot the production stack from shell scripts to code, while collaborating with networking, storage, and MSP teams.

The role requires 8–12 years of systems engineering experience, deep Linux knowledge, and hands-on work with Slurm, Kubernetes, GPFS, and cloud platforms.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff HPC Infrastructure Architect (On-Prem & Cloud)
Staff HPC Infrastructure Architect (On-Prem & Cloud)

Guardant Health, Inc. • Palo Alto (CA)

On-site
USD 173,000 - 238,000
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters

Parallel Works • Chicago (IL)

Hybrid
USD 140,000 - 190,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Senior HPC & GPU Cluster Architect
Senior HPC & GPU Cluster Architect

The Consensus • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Visa sponsorships
401(k) retirement matching
Medical, dental & vision insurance
+2
Staff HPC Infrastructure Engineer
Staff HPC Infrastructure Engineer

Guardant Health, Inc. • Palo Alto (CA)

On-site
USD 173,000 - 238,000
Senior HPC & GPU Cluster Architect — Scale & Automate
Senior HPC & GPU Cluster Architect — Scale & Automate

San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Generous equity grant
Competitive salary
Visa sponsorship
+6
HPC Linux Engineer: Kubernetes & High-Performance Infra
HPC Linux Engineer: Kubernetes & High-Performance Infra

KLA • Milpitas (CA)

On-site
USD 136,000 - 200,000
Medical, dental, vision
401(k) including company matching
Tuition reimbursement
Senior HPC Systems Administrator
Senior HPC Systems Administrator

RedLine Performance Solutions, LLC. • Berkeley (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
paid time off
401k match
health care benefits
Senior HPC Hardware Architect & Infra Automation Lead
Senior HPC Hardware Architect & Infra Automation Lead

Career Techniques • Dallas (TX)

Hybrid
USD 120,000 - 180,000
Senior HPC Systems Engineer — Slurm, GPU, Cloud-Native
Senior HPC Systems Engineer — Slurm, GPU, Cloud-Native

Nscale • New York (NY)

On-site
USD 180,000 - 260,000
Bonus
Equity
Medical Insurance
+5
Senior HPC Systems Engineer: Scale AI Clusters
Senior HPC Systems Engineer: Scale AI Clusters

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Health, vision and dental coverage
401(k) retirement plan
+2