An application made for this job — a tailored resume and cover letter that speak straight to the posting.
SpaceXAI is seeking a Site Reliability Engineer focused on campus reliability to design monitoring, lead incident command, and align compute, network, storage, power, and cooling. You will own playbooks, drive blameless postmortems, define error budgets, and collaborate with NOC and facilities teams.
The role requires 5+ years in SRE or related fields, hands-on leadership, and strong scripting in Python/Bash with experience in at least one systems language.
SpaceXAI is seeking a Site Reliability Engineer focused on campus reliability to design monitoring, lead incident command, and align compute, network, storage, power, and cooling. You will own playbooks, drive blameless postmortems, define error budgets, and collaborate with NOC and facilities teams.
The role requires 5+ years in SRE or related fields, hands-on leadership, and strong scripting in Python/Bash with experience in at least one systems language.