AI Infrastructure Engineer – GPU, Kubernetes & On-Prem

United States Digital Space LLC

United States

Remote

USD 125,000 - 195,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Stock options
401(k) plan
Medical coverage
Paid vacation
13+ holidays

Job summary

United States Digital Space LLC is hiring a Software Engineer for AI infrastructure focused on Starshield’s software and GPU platforms. You will design, deploy, and scale on-premises infrastructure to support national security missions and work across SRE, DevOps, and GPU teams.

You’ll automate deployments, support bare-metal and virtualized GPU services, and collaborate with AI engineers to build scalable, reliable products for sensitive environments.

Qualifications

  • Bachelor’s degree in computer science, information systems/IT, or an engineering discipline and 1+ years of professional experience in site reliability engineering or DevOps; OR 3+ years of professional experience in site reliability engineering or DevOps in lieu of a degree
  • 1+ years of professional experience with Linux operating systems
  • Experience with Terraform, Ansible, or other infrastructure tools
  • Experience with containerization technologies (i.e. OCI containers, Kubernetes)
  • Experience scripting in Bash, Python, or other similar languages
  • Development experience in Python, C++, or Go

Responsibilities

  • Manage GPU/CPU infrastructure deployments to Top Secret data centers
  • Manage and provide support for GPU as a service for external customers on bare metal hardware and virtualized platforms
  • Design, validate, and productize solutions for AI clusters (100k+ GPU scale)
  • Develop automation to deploy and manage on-premise Kubernetes\AI clusters, and operating systems
  • Deploy and manage core infrastructure such as databases, monitoring and distributed storage
  • Closely collaborate with AI engineers to create highly scalable, operable, and maintainable products
  • Engage in and improve the whole lifecycle of services -- from inception and design, through deployment, operation and refinement
  • Monitoring and alerting supporting systems to have high availability
  • Identify areas for improvement and create innovative solutions that enable high system availability

Skills

Linux
Python
C++
Go
Bash
DevOps

Education

Bachelor’s degree in CS/IT/Engineering

Tools

Terraform
Ansible
Kubernetes
OCI containers

Job description

United States Digital Space LLC is hiring a Software Engineer for AI infrastructure focused on Starshield’s software and GPU platforms. You will design, deploy, and scale on-premises infrastructure to support national security missions and work across SRE, DevOps, and GPU teams.

You’ll automate deployments, support bare-metal and virtualized GPU services, and collaborate with AI engineers to build scalable, reliable products for sensitive environments.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Infra Engineer - GPU, Kubernetes & On-Prem
Senior AI Infra Engineer - GPU, Kubernetes & On-Prem

United States Digital Space LLC • United States

Remote
USD 165,000 - 265,000
Stock options
Long-term incentives
Discretionary bonuses
+7
Senior AI Infrastructure Engineer - GPU & On-Prem Kubernetes
Senior AI Infrastructure Engineer - GPU & On-Prem Kubernetes

United States Digital Space LLC • El Segundo (CA), Washington

On-site
USD 165,000 - 265,000
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes
AI Infrastructure Engineer: GPU, On‑Prem & Kubernetes

Engg • Redmond (WA)

On-site
USD 125,000 - 200,000
Stock options
Medical, vision & dental coverage
401(k) plan
+2
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

SpaceX • Washington

On-site
USD 125,000 - 190,000
Stock options and long‑term incentives
Discretionary bonuses
Comprehensive medical, vision, dental
AI Infrastructure Engineer – GPU & On-Prem
AI Infrastructure Engineer – GPU & On-Prem

InvestedintheMission • Hawthorne (CA)

On-site
USD 125,000 - 195,000
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

InvestedintheMission • Redmond (WA)

On-site
USD 125,000 - 200,000
Company stock
401(k) plan
Paid time off
AI Infrastructure Engineer — GPU & On-Prem Kubernetes
AI Infrastructure Engineer — GPU & On-Prem Kubernetes

Spacex • Redmond (WA)

On-site
USD 125,000 - 200,000
Long‑term incentives
Stock options
Bonuses
+5
AI Infrastructure Engineer — GPU & On‑Prem Systems
AI Infrastructure Engineer — GPU & On‑Prem Systems

InvestedintheMission • Palo Alto (CA)

On-site
USD 125,000 - 195,000
Stock options
Health benefits
401(k)
+1
AI Infrastructure Engineer - GPU & On-Prem
AI Infrastructure Engineer - GPU & On-Prem

InvestedintheMission • Washington

On-site
USD 125,000 - 195,000
Stock options
Bonuses
Medical, vision, dental coverage
+3
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem
Senior AI Infrastructure Engineer — GPU, Kubernetes, On-Prem

Engg • Redmond (WA)

On-site
USD 165,000 - 270,000
Medical, vision, dental coverage
401(k) with company match
Paid vacation and holidays
+1