GNC Site Reliability Engineer — Scalable HPC & Tools
SPACE EXPLORATION TECHNOLOGIES CORP
Hawthorne (CA)
On-site
USD 125,000 - 175,000
Full time
14 days+
Application generator
An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Get past ATS filters
Benefits offered by this job
Medical, vision, and dental coverage
401(k) retirement plan
Paid parental leave
Discretionary bonuses
Job summary
SPACE EXPLORATION TECHNOLOGIES CORP is seeking a Site Reliability Engineer in Hawthorne, California, to manage mission-critical products for Guidance, Navigation, and Control (GNC) teams. The ideal candidate possesses a degree in a relevant field or equivalent experience, with strong skills in Linux and Python. Responsibilities include deploying scalable systems, collaborating with software engineers, and maintaining HPC clusters. The role offers competitive pay and comprehensive benefits, including stock options, medical coverage, and vacation days.
Qualifications
Bachelor’s degree or 4+ years of experience with site reliability or DevOps.
Experience with Linux operating systems and Python development frameworks.
Ability and willingness to obtain a Top Secret clearance.
Responsibilities
Deploy and scale mission-critical GNC products and services.
Monitor and maintain an HPC cluster consisting of tens of thousands of CPUs.
Closely collaborate with GNC software engineers to create maintainable products.
Skills
Linux operating systems
Python
Docker
Kubernetes
Ansible
Networking knowledge of TCP/IP
Education
Bachelor’s degree in computer science, information systems/IT, engineering, math, or scientific discipline
Tools
Terraform
GPU fleets management
Job description
SPACE EXPLORATION TECHNOLOGIES CORP is seeking a Site Reliability Engineer in Hawthorne, California, to manage mission-critical products for Guidance, Navigation, and Control (GNC) teams. The ideal candidate possesses a degree in a relevant field or equivalent experience, with strong skills in Linux and Python. Responsibilities include deploying scalable systems, collaborating with software engineers, and maintaining HPC clusters. The role offers competitive pay and comprehensive benefits, including stock options, medical coverage, and vacation days.