Senior HPC Engineer for Space AI & GPU Compute

The Aerospace Corporation

Chantilly (VA)

On-site

USD 135,000 - 203,000

Full time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health care plans
401(k) matching
Relocation assistance
Education assistance
Paid time off
Telework options
Professional development

Job summary

The Aerospace Corporation seeks an experienced HPC Engineer (Site Reliability Engineer Staff III/IV) to join our Computational Services team at our El Segundo, CA or Chantilly, VA facility.

You’ll develop, implement, and optimize HPC clusters that span on-premises and cloud environments, working closely with rocket scientists and engineers to support mission-critical technical analysis for national space assets.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or equivalent experience.
  • 7+ years of Linux system administration in an enterprise HPC environment.
  • Experience with compilers, mod & sim tools, languages, COTS, GOTs; including environment modules.
  • In-depth knowledge of Linux, networking, and HPC systems.
  • Experience with Infrastructure-as-Code and GitOps.
  • Proven experience with Slurm and HPC workloads (interactive and batch).
  • Experience provisioning AI & NVIDIA GPU technologies (CUDA).

Responsibilities

  • Collaborate with scientists and engineers on mission-critical analysis.
  • Lead cross-functional teams and mentor junior engineers.
  • Design and implement HPC solutions for diverse workloads.
  • Manage on-premise and cloud-enabled HPC clusters.
  • Develop automation and IaC solutions using Clush and GitOps.

Skills

Linux administration
Slurm scheduler
GitOps
Clush
CUDA
Environment modules
Scripting
Automation tools
Kubernetes integration

Education

Bachelor’s degree in Computer Science, Engineering, or equivalent experience

Tools

AWS ParallelCluster
Open OnDemand
XDMod
Lustre

Job description

The Aerospace Corporation seeks an experienced HPC Engineer (Site Reliability Engineer Staff III/IV) to join our Computational Services team at our El Segundo, CA or Chantilly, VA facility.

You’ll develop, implement, and optimize HPC clusters that span on-premises and cloud environments, working closely with rocket scientists and engineers to support mission-critical technical analysis for national space assets.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

High-Performance Computing (HPC) Engineer
High-Performance Computing (HPC) Engineer

The Aerospace Corporation • Chantilly (VA)

On-site
USD 135,000 - 203,000
Health care plans
401(k) matching
Relocation assistance
+4
Senior HPC Systems Engineer: Scale AI Clusters
Senior HPC Systems Engineer: Scale AI Clusters

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Health, vision and dental coverage
401(k) retirement plan
+2
Senior HPC Systems Engineer – Clusters & AI Compute
Senior HPC Systems Engineer – Clusters & AI Compute

SPACE EXPLORATION TECHNOLOGIES CORP • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Medical Vision Dental
Senior SRE: AI Infrastructure & GPU Clusters
Senior SRE: AI Infrastructure & GPU Clusters

SPACE EXPLORATION TECHNOLOGIES CORP • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Stock options
Long-term cash awards
Medical, vision, dental coverage
+5
HPC Infrastructure Architect for Space Systems
HPC Infrastructure Architect for Space Systems

Vast Space Llc. • Long Beach (CA)

On-site
USD 171,000 - 243,000
Medical insurance
Dental and vision
Paid time off
+3
Senior HPC Systems Engineer – Linux Clusters, GPUs & AI
Senior HPC Systems Engineer – Linux Clusters, GPUs & AI

SpaceX • Town of Texas (WI)

On-site
USD 140,000 - 190,000
Senior HPC Systems Engineer for High-Performance Clusters
Senior HPC Systems Engineer for High-Performance Clusters

SPACE EXPLORATION TECHNOLOGIES CORP • United States

On-site
USD 140,000 - 210,000
Senior HPC Systems Engineer: Linux Clusters & AI Readiness
Senior HPC Systems Engineer: Linux Clusters & AI Readiness

Future Ventures • Town of Texas (WI)

On-site
USD 140,000 - 210,000
Senior HPC Engineer - Linux, Kubernetes & GPUs
Senior HPC Engineer - Linux, Kubernetes & GPUs

SpaceX • Brownsville (TX)

On-site
USD 140,000 - 200,000
Senior HPC & GPU Systems Engineer – TS/SCI Onsite
Senior HPC & GPU Systems Engineer – TS/SCI Onsite

VMD Corp • Bethesda (MD)

On-site
USD 140,000 - 180,000