AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes

Engg

Redmond (WA)

On-site

USD 125,000 - 200,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Stock options
Medical, vision & dental coverage
401(k) plan
Paid vacation & holidays
Shuttle service to Redmond office

Job summary

SpaceX is seeking a Software Engineer for AI Infrastructure (Starshield) to design, operate and scale our on‑prem GPU/CPU infrastructure supporting critical national security missions. The role coversGPU deployments to Top Secret datacenters, AI clusters, Kubernetes/OST automation, and close collaboration with AI engineering teams.

You will contribute to scalable software products, high availability systems, and the lifecycle from design to operation, with opportunities for advanced security

Qualifications

  • Bachelor's degree in computer science, information systems/IT, or an engineering discipline
  • 1+ years of professional experience in site reliability engineering or DevOps or 3+ years in lieu of a degree
  • 1+ years of professional experience with Linux operating systems
  • Experience with Terraform, Ansible, or other infrastructure tools
  • Experience with containerization technologies (i.e. OCI containers, Kubernetes)
  • Development experience in Python, C++, or Go
  • Scripting in Bash, Python, or other languages
  • Knowledge of networking concepts and cloud virtualization

Responsibilities

  • Manage GPU/CPU infrastructure deployments to Top Secret datacenters
  • Provide support for GPU as a service on bare metal and virtualized platforms
  • Design, validate, and productize AI cluster solutions at large scale (100k+ GPUs)
  • Develop automation to deploy and manage on-premise Kubernetes/AI clusters and OSs
  • Deploy and manage core infrastructure such as databases, monitoring and distributed storage
  • Collaborate with AI engineers to create scalable, operable products
  • Own lifecycle from inception to deployment, operation and refinement
  • Ensure monitoring and alerting supports high availability
  • Identify improvement areas and create innovative scalable solutions

Skills

Linux
Python
C++
Go
Bash
DevOps
Networking

Education

Bachelor's degree in CS/Engineering/IT

Tools

Terraform
Ansible
Kubernetes

Job description

SpaceX is seeking a Software Engineer for AI Infrastructure (Starshield) to design, operate and scale our on‑prem GPU/CPU infrastructure supporting critical national security missions. The role coversGPU deployments to Top Secret datacenters, AI clusters, Kubernetes/OST automation, and close collaboration with AI engineering teams.

You will contribute to scalable software products, high availability systems, and the lifecycle from design to operation, with opportunities for advanced security

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

SpaceX • Washington

On-site
USD 125,000 - 190,000
Stock options and long‑term incentives
Discretionary bonuses
Comprehensive medical, vision, dental
AI Infrastructure Engineer - GPU & On-Prem
AI Infrastructure Engineer - GPU & On-Prem

InvestedintheMission • Washington

On-site
USD 125,000 - 195,000
Stock options
Bonuses
Medical, vision, dental coverage
+3
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 125,000 - 200,000
Company stock
401(k) plan
Paid time off
Senior AI Infrastructure Engineer – GPU & Kubernetes
Senior AI Infrastructure Engineer – GPU & Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
Paid vacation
Paid holidays
+2
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

Spacex • Redmond (WA)

On-site
USD 125,000 - 200,000
Long‑term incentives
Stock options
Bonuses
+5
Senior AI Infrastructure Engineer — GPU & Kubernetes
Senior AI Infrastructure Engineer — GPU & Kubernetes

Engg • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Stock options
Long-term incentives
401(k) retirement plan
+1
AI Infrastructure Engineer – GPU & On-Prem
AI Infrastructure Engineer – GPU & On-Prem

InvestedintheMission • Hawthorne (CA)

On-site
USD 125,000 - 195,000
AI Infrastructure Engineer — GPU Clusters & SRE
AI Infrastructure Engineer — GPU Clusters & SRE

SpaceX • Palo Alto (CA)

On-site
USD 125,000 - 195,000
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem

Spacex • Northern (KY)

Hybrid
USD 165,000 - 265,000
Stock options
401(k) retirement plan
Discretionary bonuses
+7
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters

Xapply • Washington

On-site
USD 165,000 - 265,000
Stock options
Medical coverage
401(k)
+1