Senior AI Infra Engineer: GPU & Kubernetes Scale Leader

Engg

Palo Alto (CA)

On-site

USD 165,000 - 265,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health insurance
Dental coverage
Vision coverage
401(k) retirement plan
Paid parental leave
Paid vacation
Paid holidays
Disability insurance
Stock options
Employee Stock Purchase Plan

Job summary

SpaceX seeks a Sr. Software Engineer, AI Infrastructure (Starshield) to design and operate the software and GPU infrastructure supporting critical national-security missions.

You will deploy on-premise compute resources, automate Kubernetes/AI clusters, and collaborate with AI teams to deliver scalable, maintainable systems. The role requires strong Linux, Kubernetes, and Python/C++/Go development, with security clearance and willingness to travel.

Qualifications

  • Bachelor’s degree in computer science, information systems/IT, or an engineering discipline and 5+ years of Linux experience or 7+ years in lieu of degree.
  • 5+ years of Kubernetes experience.
  • Experience with Terraform, Ansible, or similar infrastructure tools.
  • Experience with containerization (OCI/Kubernetes) and scripting (Bash, Python, or Go).
  • Development experience in Python, C++, or Go.

Responsibilities

  • Manage GPU/CPU infrastructure deployments to Top Secret data centers.
  • Provide GPU-as-a-service support for external customers on bare metal and virtualized platforms.
  • Design, validate, and productize AI cluster solutions (100k+ GPU scale).
  • Develop automation to deploy and manage on-premise Kubernetes/AI clusters.
  • Deploy and manage databases, monitoring, and distributed storage.
  • Collaborate with AI engineers to build scalable, operable products.
  • Engage in full lifecycle of services from design to operation and refinement.
  • Ensure high availability through monitoring and alerting.
  • Mentor junior engineers; lead the team to technical excellence.

Skills

Kubernetes
Linux
Python
C++
Go
Terraform
Ansible
Bash
CI/CD
Networking

Education

Bachelor’s degree in computer science, information systems/IT, or an engineering discipline

Tools

Terraform
Ansible
Kubernetes management

Job description

SpaceX seeks a Sr. Software Engineer, AI Infrastructure (Starshield) to design and operate the software and GPU infrastructure supporting critical national-security missions.

You will deploy on-premise compute resources, automate Kubernetes/AI clusters, and collaborate with AI teams to deliver scalable, maintainable systems. The role requires strong Linux, Kubernetes, and Python/C++/Go development, with security clearance and willingness to travel.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Infra Engineer - GPU & Kubernetes
Senior AI Infra Engineer - GPU & Kubernetes

SpaceX • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Senior AI Infrastructure Engineer – GPU & Kubernetes
Senior AI Infrastructure Engineer – GPU & Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
Paid vacation
Paid holidays
+2
Senior AI Infrastructure Engineer (GPU/Kubernetes)
Senior AI Infrastructure Engineer (GPU/Kubernetes)

Spacex • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
Medical/vision/dental
401(k) retirement plan
+3
Senior AI Infrastructure Engineer — GPU & Kubernetes
Senior AI Infrastructure Engineer — GPU & Kubernetes

Engg • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Stock options
Long-term incentives
401(k) retirement plan
+1
Senior AI Infra Engineer: GPU, On-Prem & Kubernetes
Senior AI Infra Engineer: GPU, On-Prem & Kubernetes

InvestedintheMission • Palo Alto (CA)

On-site
USD 165,000 - 265,000
Medical benefits
401(k) plan
Paid parental leave
+1
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem

Engg • Redmond (WA)

On-site
USD 165,000 - 270,000
Medical, vision, dental coverage
401(k) with company match
Paid vacation and holidays
+1
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Senior AI Infra Engineer: GPU & On-Prem Kubernetes
Senior AI Infra Engineer: GPU & On-Prem Kubernetes

Linuxconfig • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
401(k) retirement plan
Medical, vision, dental coverage
+1
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters

Xapply • Washington

On-site
USD 165,000 - 265,000
Stock options
Medical coverage
401(k)
+1
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem

Spacex • Northern (KY)

Hybrid
USD 165,000 - 265,000
Stock options
401(k) retirement plan
Discretionary bonuses
+7