AI Infra Engineer: GPU Clusters & On‑Prem Systems

Spacex

Hawthorne, Northern (CA, KY)

Hybrid

USD 125,000 - 195,000

Full time

34 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Stock options
Medical, vision, dental
401(k)
Paid time off

Job summary

SpaceX in Hawthorne, CA is seeking a Software Engineer for AI Infrastructure (Starshield). You’ll design, operate, and scale the on-prem GPU and AI infrastructure to support critical national security missions, collaborating with AI engineers and cross-functional teams.

Focus areas include automation, Kubernetes clusters, Terraform/Ansible, and performance optimization. Requires 1+ years SRE/DevOps or 3+ years without a degree; Top Secret clearance is a plus.

Qualifications

  • Bachelor’s degree in computer science, information systems/IT, or an engineering discipline and 1+ years of professional experience in site reliability engineering or DevOps; OR 3+ years of professional experience in site reliability engineering or DevOps in lieu of a degree
  • 1+ years of professional experience with Linux operating systems
  • Experience with Terraform, Ansible, or other infrastructure tools
  • Experience with containerization technologies (i.e. OCI containers, Kubernetes)
  • Experience scripting in Bash, Python, or other similar languages
  • Development experience in Python, C++, or Go

Responsibilities

  • Manage and provide support for GPU as a service for external customers on bare metal hardware and virtualized platforms
  • Design, validate, and productize solutions for AI clusters (100k+ GPU scale)
  • Develop automation to deploy and manage on-premise Kubernetes\AI clusters, and operating systems
  • Deploy and manage core infrastructure such as databases, monitoring and distributed storage
  • Closely collaborate with AI engineers to create highly scalable, operable, and maintainable products
  • Engage in and improve the whole lifecycle of services -- from inception and design, through deployment, operation and refinement
  • Monitoring and alerting supporting systems to have high availability
  • Identify areas for improvement and create innovative solutions that enable high system availability

Skills

Linux
Terraform
Ansible
Kubernetes
Python
C++
Go
Bash scripting

Education

Bachelor’s degree in computer science, information systems/IT, or an engineering discipline

Tools

OCI containers
Kubernetes
Terraform
Ansible

Job description

SpaceX in Hawthorne, CA is seeking a Software Engineer for AI Infrastructure (Starshield). You’ll design, operate, and scale the on-prem GPU and AI infrastructure to support critical national security missions, collaborating with AI engineers and cross-functional teams.

Focus areas include automation, Kubernetes clusters, Terraform/Ansible, and performance optimization. Requires 1+ years SRE/DevOps or 3+ years without a degree; Top Secret clearance is a plus.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Infra Engineer: GPU-Driven On-Prem AI Clusters
Senior AI Infra Engineer: GPU-Driven On-Prem AI Clusters

InvestedintheMission • Washington

On-site
USD 165,000 - 265,000
Stock options
Comprehensive medical/vision/dental
Paid time off and holidays
AI Infra Engineer: Scale GPU Clusters & On-Prem AI
AI Infra Engineer: Scale GPU Clusters & On-Prem AI

Xapply • Palo Alto (CA)

On-site
USD 125,000 - 195,000
Stock options
Long-term incentives
Medical, vision & dental
+5
AI Infrastructure Engineer — GPU & On‑Prem Systems
AI Infrastructure Engineer — GPU & On‑Prem Systems

InvestedintheMission • Palo Alto (CA)

On-site
USD 125,000 - 195,000
Stock options
Health benefits
401(k)
+1
AI Infrastructure Engineer — GPU Clusters & SRE
AI Infrastructure Engineer — GPU Clusters & SRE

SpaceX • Palo Alto (CA)

On-site
USD 125,000 - 195,000
AI Infrastructure Engineer – GPU & On-Prem
AI Infrastructure Engineer – GPU & On-Prem

InvestedintheMission • Hawthorne (CA)

On-site
USD 125,000 - 195,000
AI Infrastructure Engineer – GPU & On-Prem Cloud
AI Infrastructure Engineer – GPU & On-Prem Cloud

Linuxconfig • Hawthorne (CA), Northern (KY)

Hybrid
USD 125,000 - 195,000
Stock options
Medical, vision & dental coverage
401(k) retirement plan
+1
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters
Senior AI Infrastructure Engineer — GPU & On-Prem Clusters

Xapply • Washington

On-site
USD 165,000 - 265,000
Stock options
Medical coverage
401(k)
+1
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation
Senior AI Infra Engineer: GPU Clusters, Kubernetes & Automation

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem
Senior AI Infrastructure Engineer - GPU Clusters & On-Prem

Spacex • Northern (KY)

Hybrid
USD 165,000 - 265,000
Stock options
401(k) retirement plan
Discretionary bonuses
+7
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes

Engg • Redmond (WA)

On-site
USD 125,000 - 200,000
Stock options
Medical, vision & dental coverage
401(k) plan
+2