Senior HPC Systems Engineer: Linux Clusters, AI & GPU

United States Digital Space LLC

Starbase (TX)

On-site

USD 120,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

United States Digital Space LLC seeks an Sr. HPC Systems Engineer to administrate HPC clusters, storage, and fast networks. You will support engineers, install Linux compute clusters, and author clear documentation for complex systems.

Ideal candidates have 5+ years in HPC or 7+ years software experience with a degree, plus Kubernetes and container tooling proficiency. This role involves extended hours as needed.

Qualifications

  • Bachelor's degree in computer science, engineering, math, or scientific discipline and 5+ years of systems engineering experience; OR 7+ years of professional experience building software in lieu of a degree
  • 5+ years of hands-on experience with client and server hardware/software, management tools, enterprise networking, virtualization, and security technologies
  • Experience with Kubernetes

Responsibilities

  • Administer and manage HPC clusters, storage systems, and high-speed networks
  • Provide application support to the company employees across engineering disciplines
  • Install and integrate Linux-based compute clusters
  • Write instructional documentation and convey highly technical ideas in non-technical terms

Skills

Kubernetes
Linux systems
Scripting Bash Python
Docker Podman Singularity
HPC clusters
Monitoring Prometheus Grafana Nagios
Cluster schedulers Slurm PBS LSF
Automation Puppet Ansible
Virtualization
Networking security
GPU CUDA

Education

Bachelor's degree in computer science, engineering, math, or scientific discipline

Tools

Docker
Podman
Singularity
Kubernetes
Puppet
Ansible
Prometheus
Grafana
Nagios
Slurm
PBS
LSF

Job description

United States Digital Space LLC seeks an Sr. HPC Systems Engineer to administrate HPC clusters, storage, and fast networks. You will support engineers, install Linux compute clusters, and author clear documentation for complex systems.

Ideal candidates have 5+ years in HPC or 7+ years software experience with a degree, plus Kubernetes and container tooling proficiency. This role involves extended hours as needed.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Systems Architect for AI & GPU Clusters
HPC Systems Architect for AI & GPU Clusters

United States Digital Space LLC • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Senior HPC Systems Engineer: Linux Clusters & AI Readiness
Senior HPC Systems Engineer: Linux Clusters & AI Readiness

Future Ventures • Town of Texas (WI)

On-site
USD 140,000 - 210,000
Senior HPC Systems Engineer: Linux Clusters & AI Training
Senior HPC Systems Engineer: Linux Clusters & AI Training

SpaceX • Pennsylvania

On-site
USD 150,000 - 190,000
Senior HPC Systems Engineer: Scale AI Clusters
Senior HPC Systems Engineer: Scale AI Clusters

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Health, vision and dental coverage
401(k) retirement plan
+2
Senior HPC Systems Engineer – Clusters & AI Compute
Senior HPC Systems Engineer – Clusters & AI Compute

SPACE EXPLORATION TECHNOLOGIES CORP • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Medical Vision Dental
Senior HPC Systems Engineer - AI Compute & Clusters
Senior HPC Systems Engineer - AI Compute & Clusters

InvestedintheMission • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Long-term incentives
Bonuses
+6
Senior HPC Systems Engineer: High-Performance Compute & AI
Senior HPC Systems Engineer: High-Performance Compute & AI

InvestedintheMission • Town of Texas (WI)

On-site
USD 120,000 - 190,000
Senior HPC Systems Engineer for High-Performance Clusters
Senior HPC Systems Engineer for High-Performance Clusters

SPACE EXPLORATION TECHNOLOGIES CORP • United States

On-site
USD 140,000 - 210,000
Senior HPC Engineer - Linux, Kubernetes & GPUs
Senior HPC Engineer - Linux, Kubernetes & GPUs

SpaceX • Brownsville (TX)

On-site
USD 140,000 - 200,000
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters

Parallel Works • Chicago (IL)

Hybrid
USD 140,000 - 190,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1