Lead HPC Engineer — On-Site, Slurm & GPU Expert

The Aerospace Corporation

United States

On-site

USD 135,000 - 203,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Comprehensive health care
401(k) plan
Relocation assistance
Education assistance

Job summary

The Aerospace Corporation seeks a High-Performance Computing (HPC) Engineer (Site Reliability Engineer Staff III/IV) to join our Computational Services team. You will design, implement, and optimize on-premises and cloud HPC clusters while collaborating with scientists and engineers on mission-critical analysis.

The role requires advanced Linux administration, Slurm knowledge, and automation experience; TS/SCI clearance is preferred.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or equivalent experience.
  • 7+ years Linux system administration in an enterprise HPC environment.
  • Experience with Slurm scheduler and environment modules.
  • Experience with Infrastructure-as-Code and GitOps practices.
  • Experience with AI & GPU technologies (CUDA) and scripting.

Responsibilities

  • Collaborate with scientists and engineers on mission-critical technical analysis for national space assets.
  • Lead cross-functional teams and mentor junior engineers.
  • Design and optimize HPC clusters for on-premises and cloud environments.
  • Manage 10,000-core classified and 5,000-core unclassified clusters.
  • Develop automation using tools like Clush and implement IaC and GitOps.
  • Harden Linux systems to meet security requirements.

Skills

Linux administration
Slurm scheduler
Automation scripting
GitOps
Clush
HPC design
Communication skills
Security+ / DoD 8570

Education

Bachelor’s degree in Computer Science, Engineering, or equivalent experience

Tools

Environment modules
CUDA (NVIDIA)
Ansible
Terraform
Open OnDemand

Job description

The Aerospace Corporation seeks a High-Performance Computing (HPC) Engineer (Site Reliability Engineer Staff III/IV) to join our Computational Services team. You will design, implement, and optimize on-premises and cloud HPC clusters while collaborating with scientists and engineers on mission-critical analysis.

The role requires advanced Linux administration, Slurm knowledge, and automation experience; TS/SCI clearance is preferred.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Space HPC Engineer & SRE - On-Site, TS/SCI Ready
Space HPC Engineer & SRE - On-Site, TS/SCI Ready

The Aerospace Corporation • El Segundo (CA)

On-site
USD 135,000 - 203,000
Health care benefits
Paid time off
Telework options
Senior HPC Engineer for Space AI & GPU Compute
Senior HPC Engineer for Space AI & GPU Compute

The Aerospace Corporation • Chantilly (VA)

On-site
USD 135,000 - 203,000
Health care plans
401(k) matching
Relocation assistance
+4
Senior HPC & GPU Systems Engineer – TS/SCI Onsite
Senior HPC & GPU Systems Engineer – TS/SCI Onsite

VMD Corp • Bethesda (MD)

On-site
USD 140,000 - 180,000
HPC Cloud Engineer - GPU & Slurm Cluster Specialist
HPC Cloud Engineer - GPU & Slurm Cluster Specialist

Engg • Hill Air Force Base (UT)

On-site
USD 106,000 - 206,000
Senior HPC & GPU Systems Engineer — On-Site Bethesda
Senior HPC & GPU Systems Engineer — On-Site Bethesda

RPMGlobal • Bethesda (MD), Northern (KY)

Hybrid
USD 140,000 - 200,000
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux

Strategic Business Systems, Inc (SBS) • Chantilly (VA)

Hybrid
USD 120,000 - 180,000
Flexible work arrangements
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters

Parallel Works • Chicago (IL)

Hybrid
USD 140,000 - 190,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Senior HPC & GPU Systems Engineer - On‑Site (Top Secret)
Senior HPC & GPU Systems Engineer - On‑Site (Top Secret)

RPMGlobal • Bethesda (MD), Northern (KY)

Hybrid
USD 110,000 - 160,000
High-Performance Computing (HPC) Engineer
High-Performance Computing (HPC) Engineer

The Aerospace Corporation • Chantilly (VA)

On-site
USD 135,000 - 203,000
Health care plans
401(k) matching
Relocation assistance
+4
HPC Infrastructure Architect for Space Systems
HPC Infrastructure Architect for Space Systems

Vast Space Llc. • Long Beach (CA)

On-site
USD 171,000 - 243,000
Medical insurance
Dental and vision
Paid time off
+3