AI Infrastructure Engineer — GPU & On-Prem Kubernetes

Spacex

Redmond (WA)

On-site

USD 125,000 - 200,000

Full time

31 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Long‑term incentives
Stock options
Bonuses
Healthcare
401(k)
Paid vacation
Paid holidays
Shuttle service

Job summary

SpaceX is seeking a Software Engineer for AI Infrastructure (Starshield) in Redmond, WA. You will design, operate, and scale GPU-focused infrastructure to support critical national security missions, including on-premise clusters and related software products.

The role involves collaborating with AI engineers, building scalable automation, and advancing Kubernetes/AI infrastructure with a focus on reliability and performance. TS clearance is preferred or required to obtain.

Qualifications

  • Bachelors degree in a technical field and 1+ years in SRE/DevOps or 3+ years in SRE/DevOps without a degree.
  • 1+ years of Linux experience.
  • Experience with Terraform, Ansible, or other infrastructure tools.
  • Experience with containerization technologies (i.e. OCI containers, Kubernetes).
  • Scripting in Bash, Python, or similar languages.
  • Development experience in Python, C++, or Go.

Responsibilities

  • Manage and provide support for GPU as a service on bare metal hardware and virtualized platforms.
  • Design, validate, and productize AI cluster solutions (100k+ GPU scale).
  • Develop automation to deploy and manage on-premise Kubernetes/AI clusters and OSes.
  • Deploy and manage core infrastructure such as databases, monitoring and storage.
  • Collaborate with AI engineers to create scalable, maintainable products.
  • Engage in the full lifecycle of services from inception to refinement.
  • Monitor and alert supporting systems for high availability.
  • Identify areas for improvement and create innovative solutions.

Skills

Linux
Terraform
Ansible
Kubernetes
Bash
Python
C++
Go
SRE/DevOps

Education

Bachelor's degree

Tools

OCI containers

Job description

SpaceX is seeking a Software Engineer for AI Infrastructure (Starshield) in Redmond, WA. You will design, operate, and scale GPU-focused infrastructure to support critical national security missions, including on-premise clusters and related software products.

The role involves collaborating with AI engineers, building scalable automation, and advancing Kubernetes/AI infrastructure with a focus on reliability and performance. TS clearance is preferred or required to obtain.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 125,000 - 200,000
Company stock
401(k) plan
Paid time off
Senior AI Infra Engineer: GPU & On-Prem Kubernetes
Senior AI Infra Engineer: GPU & On-Prem Kubernetes

Linuxconfig • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
401(k) retirement plan
Medical, vision, dental coverage
+1
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem

Engg • Redmond (WA)

On-site
USD 165,000 - 270,000
Medical, vision, dental coverage
401(k) with company match
Paid vacation and holidays
+1
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

SpaceX • Washington

On-site
USD 125,000 - 190,000
Stock options and long‑term incentives
Discretionary bonuses
Comprehensive medical, vision, dental
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes

Engg • Redmond (WA)

On-site
USD 125,000 - 200,000
Stock options
Medical, vision & dental coverage
401(k) plan
+2
AI Infrastructure Engineer – GPU & On-Prem
AI Infrastructure Engineer – GPU & On-Prem

InvestedintheMission • Hawthorne (CA)

On-site
USD 125,000 - 195,000
AI Infrastructure Engineer — GPU & On‑Prem Systems
AI Infrastructure Engineer — GPU & On‑Prem Systems

InvestedintheMission • Palo Alto (CA)

On-site
USD 125,000 - 195,000
Stock options
Health benefits
401(k)
+1
Senior AI Infrastructure Engineer (GPU/Kubernetes)
Senior AI Infrastructure Engineer (GPU/Kubernetes)

Spacex • Redmond (WA)

On-site
USD 165,000 - 270,000
Stock options
Medical/vision/dental
401(k) retirement plan
+3
AI Infrastructure SRE — GPU & On-Prem Kubernetes
AI Infrastructure SRE — GPU & On-Prem Kubernetes

SPACE EXPLORATION TECHNOLOGIES CORP • Washington

On-site
USD 125,000 - 195,000
Stock options/long-term incentives
Medical, Vision, Dental
401(k) plan
+2
AI Infrastructure SRE - GPU Clusters
AI Infrastructure SRE - GPU Clusters

SPACE EXPLORATION TECHNOLOGIES CORP • Redmond (WA), Northern (KY)

Hybrid
USD 125,000 - 200,000
Stock options
401(k) plan
Comprehensive medical benefits
+1