Senior AI Infra Engineer: GPU-Driven On-Prem AI Clusters

InvestedintheMission

Washington (District of Columbia)

On-site

USD 165,000 - 265,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Stock options
Comprehensive medical/vision/dental
Paid time off and holidays

Job summary

SpaceX is seeking a Sr. Software Engineer, AI Infrastructure (Starshield) to design, operate, and scale on‑prem and GPU-intensive infrastructure supporting national security missions.

You will deploy AI clusters, manage Kubernetes across data centers, and mentor engineers while collaborating with AI teams to deliver reliable, scalable software. You must clear Top Secret/SOI or DOE level Q, travel as needed, and align with SpaceX ITAR requirements.

Qualifications

  • Bachelor’s degree in computer science, information systems/IT, or an engineering discipline with 5+ years of Linux experience; or 7+ years in software/DevOps/SRE.
  • 5+ years Kubernetes experience.
  • 5+ years Linux OS management experience.
  • Terraform, Ansible, or similar infra tools experience.
  • Containerization tech experience (OCI containers, Kubernetes).
  • Scripting in Bash, Python, or similar languages.
  • Python, C++, or Go development experience.

Responsibilities

  • Manage GPU/CPU infra deployments to Top Secret data centers.
  • Provide GPU-as-a-service support on bare metal and virtualized platforms.
  • Design and productize AI clusters (100k+ GPU scale).
  • Develop automation to deploy and manage on‑prem Kubernetes/AI clusters and OSes.
  • Deploy and manage core infra: databases, monitoring, distributed storage.
  • Collaborate with AI engineers to deliver scalable, operable products.
  • Oversee lifecycle from design to deployment and refinement.
  • Maintain monitoring and alerting for high availability.
  • Identify improvement areas and create innovative solutions for reliability.
  • Mentor junior engineers.
  • Lead the team to technical excellence as a senior engineer.

Skills

Kubernetes
Linux
Python
Go
C++
Terraform
Ansible
Bash

Education

Bachelor’s degree in CS/Engineering

Tools

Docker
CI/CD tooling

Job description

SpaceX is seeking a Sr. Software Engineer, AI Infrastructure (Starshield) to design, operate, and scale on‑prem and GPU-intensive infrastructure supporting national security missions.

You will deploy AI clusters, manage Kubernetes across data centers, and mentor engineers while collaborating with AI teams to deliver reliable, scalable software. You must clear Top Secret/SOI or DOE level Q, travel as needed, and align with SpaceX ITAR requirements.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Infrastructure Engineer - GPU Clusters & On-Prem
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem

Spacex • Northern (KY)

Hybrid
USD 165,000 - 265,000
Stock options
401(k) retirement plan
Discretionary bonuses
+7
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters

Xapply • Washington

On-site
USD 165,000 - 265,000
Stock options
Medical coverage
401(k)
+1
Senior AI Infra Engineer - GPU & Kubernetes
Senior AI Infra Engineer - GPU & Kubernetes

SpaceX • Palo Alto (CA)

On-site
USD 165,000 - 265,000
AI Infrastructure Engineer - GPU & On-Prem
AI Infrastructure Engineer - GPU & On-Prem

InvestedintheMission • Washington

On-site
USD 125,000 - 195,000
Stock options
Bonuses
Medical, vision, dental coverage
+3
AI Infra Engineer: GPU Clusters & On‑Prem Systems
AI Infra Engineer: GPU Clusters & On‑Prem Systems

Spacex • Hawthorne (CA), Northern (KY)

Hybrid
USD 125,000 - 195,000
Stock options
Medical, vision, dental
401(k)
+1
Senior AI Infra Engineer: GPU & Kubernetes Scale Leader
Senior AI Infra Engineer: GPU & Kubernetes Scale Leader

Engg • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Health insurance
Dental coverage
Vision coverage
+7
Senior AI Infrastructure Engineer — GPU & Kubernetes
Senior AI Infrastructure Engineer — GPU & Kubernetes

Engg • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Stock options
Long-term incentives
401(k) retirement plan
+1
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Senior AI Infrastructure Engineer – GPU & Kubernetes
Senior AI Infrastructure Engineer – GPU & Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
Paid vacation
Paid holidays
+2
Senior AI Infra Engineer: GPU, On-Prem & Kubernetes
Senior AI Infra Engineer: GPU, On-Prem & Kubernetes

InvestedintheMission • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Medical benefits
401(k) plan
Paid parental leave
+1