A complete application in a minute — tailored resume and cover letter, ready to send.
SpaceX is seeking a Sr. Site Reliability Engineer for the Starshield program to design, operate, and scale on-premise GPU/CPU infrastructure and AI clusters.
You will develop automation to deploy Kubernetes and related OSes, deploy databases and monitoring, and collaborate with AI engineers to deliver reliable services. The role requires deep Linux and Kubernetes expertise, Terraform/Ansible experience, and the ability to lead a team toward technical excellence while enabling high availability
SpaceX is seeking a Sr. Site Reliability Engineer for the Starshield program to design, operate, and scale on-premise GPU/CPU infrastructure and AI clusters.
You will develop automation to deploy Kubernetes and related OSes, deploy databases and monitoring, and collaborate with AI engineers to deliver reliable services. The role requires deep Linux and Kubernetes expertise, Terraform/Ansible experience, and the ability to lead a team toward technical excellence while enabling high availability