Space HPC Engineer & SRE - On-Site, TS/SCI Ready

The Aerospace Corporation

El Segundo (CA)

On-site

USD 135,000 - 203,000

Full time

7 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health care benefits
Paid time off
Telework options

Job summary

The Aerospace Corporation in El Segundo, CA is seeking an HPC Engineer (Site Reliability Engineer Staff III/IV) to join our Computational Services team. You will design, deploy, and optimize HPC clusters on prem and in cloud, collaborating with rocket scientists and engineers on mission-critical analyses.

Minimum qualifications include a BS in CS/engineering, 7+ years Linux admin in HPC, Slurm expertise, and experience with IaC, GitOps, and AI/NVIDIA GPUs.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or equivalent experience.
  • Minimum of 7 years’ experience in Linux system administration within an enterprise HPC environment.
  • Experience with environment modules, compilers, and HPC toolchains.
  • Proficiency with Slurm scheduler management for interactive and batch workloads.
  • Proficiency in scripting and automation tools (e.g., Clush).
  • Experience with Infrastructure‑as‑Code and GitOps practices.
  • Experience with AI & NVIDIA GPU technologies (CUDA).
  • Ability to obtain and maintain TS/SCI clearance (U.S. citizenship required).

Responsibilities

  • Collaborate with scientists and engineers on mission-critical analyses for space assets.
  • Lead cross-functional teams and mentor junior engineers.
  • Design and implement HPC solutions optimizing resource use across workloads.
  • Manage on-premise and cluster performance and security.
  • Develop automation and configuration management using IaC/GitOps.
  • Support GPU-enabled workloads and Slurm deployments.

Skills

Linux system administration
Leadership
Scripting
GPU computing
Automation

Education

Bachelor's degree in Computer Science or Engineering

Tools

GitOps
Clush
Infrastructure-as-Code
CUDA/NVIDIA GPUs

Job description

The Aerospace Corporation in El Segundo, CA is seeking an HPC Engineer (Site Reliability Engineer Staff III/IV) to join our Computational Services team. You will design, deploy, and optimize HPC clusters on prem and in cloud, collaborating with rocket scientists and engineers on mission-critical analyses.

Minimum qualifications include a BS in CS/engineering, 7+ years Linux admin in HPC, Slurm expertise, and experience with IaC, GitOps, and AI/NVIDIA GPUs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Space HPC Engineer & SRE (Slurm, GPU, Cloud)
Space HPC Engineer & SRE (Slurm, GPU, Cloud)

aero • El Segundo (CA)

On-site
USD 140,000 - 190,000
High-Performance Computing (HPC) Engineer
High-Performance Computing (HPC) Engineer

aero • El Segundo (CA)

On-site
USD 140,000 - 190,000
Senior HPC Systems Engineer – Clusters & AI Compute
Senior HPC Systems Engineer – Clusters & AI Compute

SPACE EXPLORATION TECHNOLOGIES CORP • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Medical Vision Dental
High-Performance Computing (HPC) Engineer
High-Performance Computing (HPC) Engineer

The Aerospace Corporation • El Segundo (CA)

On-site
USD 135,000 - 203,000
Health care benefits
Paid time off
Telework options
Kubernetes SRE — Space-Grade PaaS & AI Workloads
Kubernetes SRE — Space-Grade PaaS & AI Workloads

The Aerospace Corporation • El Segundo (CA), Northern (KY)

Hybrid
USD 129,000 - 194,000
Comprehensive health care
401(k) plan with matching
Relocation assistance
+2
HPC Site Reliability Engineer — Automation & Infra
HPC Site Reliability Engineer — Automation & Infra

SPACE EXPLORATION TECHNOLOGIES CORP • Redmond (WA)

On-site
USD 125,000 - 150,000
Health insurance
401(k) retirement plan
Paid vacation and holidays
+1
TS/SCI Linux HPC Engineer (Onsite)
TS/SCI Linux HPC Engineer (Onsite)

Jobless • Charlottesville (VA), Northern (KY)

Hybrid
USD 100,000 - 175,000
HPC Linux Systems Engineer (TS/SCI) – Onsite
HPC Linux Systems Engineer (TS/SCI) – Onsite

Plus3 IT Systems • Charlottesville (VA)

On-site
USD 100,000 - 175,000
Health benefits
401(k) matching
Parental leave
Senior HPC Systems Engineer: Scale AI Clusters
Senior HPC Systems Engineer: Scale AI Clusters

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Health, vision and dental coverage
401(k) retirement plan
+2
Senior HPC Systems Engineer for High-Performance Clusters
Senior HPC Systems Engineer for High-Performance Clusters

SPACE EXPLORATION TECHNOLOGIES CORP • United States

On-site
USD 140,000 - 210,000