AI Infrastructure Engineer — GPU & On‑Prem Systems

InvestedintheMission

Palo Alto (CA)

On-site

USD 125,000 - 195,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Stock options
Health benefits
401(k)
Paid time off

Job summary

SpaceX in Palo Alto is seeking a Software Engineer for AI infrastructure within the Starshield program. You will design, operate, and scale on‑premise GPU and AI clusters, deploying Kubernetes, databases, and distributed storage to support national security missions.

The role emphasizes automation, collaboration with AI engineers, and building highly available services across data centers; requires TS clearance or ability to obtain, and willingness for extended hours and travel.

Qualifications

  • Bachelor’s degree in computer science, information systems/IT, or engineering with 1+ years in SRE/DevOps or 3+ years without a degree.
  • 1+ years of Linux experience and scripting (Bash, Python).
  • Experience with Terraform, Ansible, or similar infra tools; containerization with Kubernetes.
  • Development experience in Python, C++, or Go.

Responsibilities

  • Manage GPU/CPU infrastructure deployments to Top Secret data centers.
  • Provide GPU as a service for external customers on bare metal and virtualized platforms.
  • Design and productize AI cluster solutions at 100k+ GPU scale.
  • Develop automation to deploy and manage on-premise Kubernetes/AI clusters and OSes.
  • Deploy and manage databases, monitoring, and distributed storage.
  • Collaborate with AI engineers to build scalable, operable products.
  • Improve the full lifecycle: design, deployment, operation and refinement.
  • Ensure high availability with monitoring and alerting.

Skills

Linux experience
Infrastructure automation
Python
Go
C++
Bash scripting
Strong communications
Top Secret clearance

Education

Bachelor's degree in CS or engineering

Tools

Terraform
Ansible
Kubernetes

Job description

SpaceX in Palo Alto is seeking a Software Engineer for AI infrastructure within the Starshield program. You will design, operate, and scale on‑premise GPU and AI clusters, deploying Kubernetes, databases, and distributed storage to support national security missions.

The role emphasizes automation, collaboration with AI engineers, and building highly available services across data centers; requires TS clearance or ability to obtain, and willingness for extended hours and travel.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes

Engg • Redmond (WA)

On-site
USD 125,000 - 200,000
Stock options
Medical, vision & dental coverage
401(k) plan
+2
AI Infrastructure Engineer — GPU Clusters & SRE
AI Infrastructure Engineer — GPU Clusters & SRE

SpaceX • Palo Alto (CA)

On-site
USD 125,000 - 195,000
AI Infrastructure Engineer – GPU & On-Prem
AI Infrastructure Engineer – GPU & On-Prem

InvestedintheMission • Hawthorne (CA)

On-site
USD 125,000 - 195,000
AI Infrastructure Engineer - GPU & On-Prem
AI Infrastructure Engineer - GPU & On-Prem

InvestedintheMission • Washington

On-site
USD 125,000 - 195,000
Stock options
Bonuses
Medical, vision, dental coverage
+3
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

SpaceX • Washington

On-site
USD 125,000 - 190,000
Stock options and long‑term incentives
Discretionary bonuses
Comprehensive medical, vision, dental
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem

Spacex • Northern (KY)

Hybrid
USD 165,000 - 265,000
Stock options
401(k) retirement plan
Discretionary bonuses
+7
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters

Xapply • Washington

On-site
USD 165,000 - 265,000
Stock options
Medical coverage
401(k)
+1
Senior AI Infra Engineer - GPU & Kubernetes
Senior AI Infra Engineer - GPU & Kubernetes

SpaceX • Palo Alto (CA)

On-site
USD 165,000 - 265,000
AI Infrastructure Engineer – GPU & On-Prem Cloud
AI Infrastructure Engineer – GPU & On-Prem Cloud

Linuxconfig • Hawthorne (CA), Northern (KY)

Hybrid
USD 125,000 - 195,000
Stock options
Medical, vision & dental coverage
401(k) retirement plan
+1
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

Spacex • Redmond (WA)

On-site
USD 125,000 - 200,000
Long‑term incentives
Stock options
Bonuses
+5