Senior AI Infra Engineer: GPU, On-Prem & Kubernetes

InvestedintheMission

Palo Alto (CA)

On-site

USD 165,000 - 265,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Medical benefits
401(k) plan
Paid parental leave
Vacation days

Job summary

SpaceX is seeking a Sr. Software Engineer, AI Infrastructure (Starshield) in Palo Alto to design, operate and scale on-premise GPU/CPU infrastructure for a national security satellite constellation.

You’ll automate deployments, build scalable Kubernetes AI clusters, and collaborate across engineering teams. The role requires 5+ years in Linux/SRE/DevOps, strong Python/C++/Go development, and experience with Terraform/Ansible.

Qualifications

  • Bachelor’s degree in computer science, information systems/IT, or an engineering discipline with 5+ years of Linux OS experience, or 7+ years in software/DevOps/SRE in lieu of a degree.
  • 5+ years of experience with Kubernetes
  • 5+ years of experience managing Linux operating systems
  • Experience with Terraform, Ansible, or other infrastructure tools
  • Experience with containerization technologies (OCI containers, Kubernetes)
  • Experience scripting in Bash, Python, or similar languages
  • Development experience in Python, C++, or Go

Responsibilities

  • Manage GPU/CPU infrastructure deployments to Top Secret data centers
  • Provide support for GPU as a service on bare metal hardware and virtualized platforms
  • Design, validate, and productize AI clusters (100k+ GPU scale)
  • Develop automation to deploy and manage on-premise Kubernetes/AI clusters and operating systems
  • Deploy and manage core infrastructure such as databases, monitoring and distributed storage
  • Collaborate with AI engineers to create scalable, operable products
  • Oversee the full lifecycle of services from design to deployment and refinement
  • Implement monitoring and alerting for high availability
  • Identify improvement areas and create innovative solutions for availability
  • Mentor and train junior engineers
  • Lead the team to technical excellence and guide decisions

Skills

Kubernetes
Linux
Python
C++
Go
Bash scripting
Python development

Education

Bachelor’s degree in CS/IT or related field

Tools

Terraform
Ansible
Kubernetes

Job description

SpaceX is seeking a Sr. Software Engineer, AI Infrastructure (Starshield) in Palo Alto to design, operate and scale on-premise GPU/CPU infrastructure for a national security satellite constellation.

You’ll automate deployments, build scalable Kubernetes AI clusters, and collaborate across engineering teams. The role requires 5+ years in Linux/SRE/DevOps, strong Python/C++/Go development, and experience with Terraform/Ansible.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Infrastructure Engineer — GPU & Kubernetes
Senior AI Infrastructure Engineer — GPU & Kubernetes

Engg • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Stock options
Long-term incentives
401(k) retirement plan
+1
Senior AI Infra Engineer: GPU & Kubernetes Scale Leader
Senior AI Infra Engineer: GPU & Kubernetes Scale Leader

Engg • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Health insurance
Dental coverage
Vision coverage
+7
Senior AI Infra Engineer - GPU & Kubernetes
Senior AI Infra Engineer - GPU & Kubernetes

SpaceX • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Senior AI Infra Engineer: GPU & On-Prem Kubernetes
Senior AI Infra Engineer: GPU & On-Prem Kubernetes

Linuxconfig • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
401(k) retirement plan
Medical, vision, dental coverage
+1
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters

Xapply • Washington

On-site
USD 165,000 - 265,000
Stock options
Medical coverage
401(k)
+1
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem

Engg • Redmond (WA)

On-site
USD 165,000 - 270,000
Medical, vision, dental coverage
401(k) with company match
Paid vacation and holidays
+1
Senior AI Infrastructure Engineer – GPU & Kubernetes
Senior AI Infrastructure Engineer – GPU & Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
Paid vacation
Paid holidays
+2
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem

Spacex • Northern (KY)

Hybrid
USD 165,000 - 265,000
Stock options
401(k) retirement plan
Discretionary bonuses
+7
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

SpaceX • Washington

On-site
USD 125,000 - 190,000
Stock options and long‑term incentives
Discretionary bonuses
Comprehensive medical, vision, dental
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes

Engg • Redmond (WA)

On-site
USD 125,000 - 200,000
Stock options
Medical, vision & dental coverage
401(k) plan
+2