Senior AI Infrastructure Engineer — GPU & On-Prem Clusters

Xapply

Washington (District of Columbia)

On-site

USD 165,000 - 265,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Stock options
Medical coverage
401(k)
Paid time off

Job summary

SpaceX is seeking a Senior Software Engineer for AI Infrastructure (Starshield) to design, operate, and scale GPU/CPU infrastructures supporting critical national security missions. You will deploy on-premise resources, build scalable software, and collaborate across teams to deliver reliable AI clusters.

The role emphasizes Kubernetes, Terraform/Ansible automation, and developing maintainable systems while mentoring junior engineers and guiding technical excellence.

Qualifications

  • Bachelor’s degree in computer science, information systems/IT, or an engineering discipline
  • 5+ years of professional experience with Linux operating systems or 7+ years in software/DevOps in lieu of a degree
  • 5+ years of experience with Kubernetes
  • Experience with Terraform, Ansible, or other infrastructure tools
  • Experience with containerization technologies (OCI containers, Kubernetes)
  • Experience scripting in Bash, Python, or similar languages
  • Development experience in Python, C++, or Go

Responsibilities

  • Manage GPU/CPU infrastructure deployments to Top Secret data centers
  • Provide support for GPU as a service on bare metal hardware and virtualized platforms
  • Design, validate, and productize AI cluster solutions (100k+ GPUs)
  • Develop automation to deploy and manage on-premise Kubernetes/AI clusters and OSes
  • Deploy and manage core infrastructure such as databases, monitoring and distributed storage
  • Collaborate with AI engineers to create scalable, operable products
  • Oversee lifecycles from design to deployment and refinement
  • Maintain high availability through monitoring and alerting
  • Identify improvements and create innovative high-availability solutions
  • Mentor and train junior engineers
  • Lead the team to technical excellence as a senior engineer

Skills

Kubernetes
Linux
Python
Go
C++
Bash

Education

Bachelor's degree in CS/IT/Engineering

Tools

Terraform
Ansible
Docker
Bazel
Makefiles

Job description

SpaceX is seeking a Senior Software Engineer for AI Infrastructure (Starshield) to design, operate, and scale GPU/CPU infrastructures supporting critical national security missions. You will deploy on-premise resources, build scalable software, and collaborate across teams to deliver reliable AI clusters.

The role emphasizes Kubernetes, Terraform/Ansible automation, and developing maintainable systems while mentoring junior engineers and guiding technical excellence.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Infrastructure Engineer - GPU Clusters & On-Prem
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem

Spacex • Northern (KY)

Hybrid
USD 165,000 - 265,000
Stock options
401(k) retirement plan
Discretionary bonuses
+7
Senior AI Infrastructure Engineer — GPU & Kubernetes
Senior AI Infrastructure Engineer — GPU & Kubernetes

Engg • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Stock options
Long-term incentives
401(k) retirement plan
+1
Senior AI Infrastructure Engineer – GPU & Kubernetes
Senior AI Infrastructure Engineer – GPU & Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
Paid vacation
Paid holidays
+2
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Senior AI Infra Engineer: GPU-Driven On-Prem AI Clusters
Senior AI Infra Engineer: GPU-Driven On-Prem AI Clusters

InvestedintheMission • Washington

On-site
USD 165,000 - 265,000
Stock options
Comprehensive medical/vision/dental
Paid time off and holidays
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem

Engg • Redmond (WA)

On-site
USD 165,000 - 270,000
Medical, vision, dental coverage
401(k) with company match
Paid vacation and holidays
+1
Senior AI Infra Engineer: GPU, On-Prem & Kubernetes
Senior AI Infra Engineer: GPU, On-Prem & Kubernetes

InvestedintheMission • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Medical benefits
401(k) plan
Paid parental leave
+1
AI Infrastructure Engineer — GPU & On‑Prem Systems
AI Infrastructure Engineer — GPU & On‑Prem Systems

InvestedintheMission • Palo Alto (CA)

On-site
USD 125,000 - 195,000
Stock options
Health benefits
401(k)
+1
Senior AI Infra Engineer: GPU & Kubernetes Scale Leader
Senior AI Infra Engineer: GPU & Kubernetes Scale Leader

Engg • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Health insurance
Dental coverage
Vision coverage
+7
Senior AI Infra Engineer - GPU & Kubernetes
Senior AI Infra Engineer - GPU & Kubernetes

SpaceX • Palo Alto (CA)

On-site
USD 165,000 - 265,000