An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Thinking Machines is hiring a Site Reliability Engineer to drive the reliability of Tinker end-to-end. You will define SLAs, build observability, and lead incident response across the platform, coordinating with security and research teams.
The role requires strong distributed-systems experience and Kubernetes expertise in a fast-growing AI infrastructure environment. Located in San Francisco, this role offers competitive compensation and visa sponsorship.
Thinking Machines is hiring a Site Reliability Engineer to drive the reliability of Tinker end-to-end. You will define SLAs, build observability, and lead incident response across the platform, coordinating with security and research teams.
The role requires strong distributed-systems experience and Kubernetes expertise in a fast-growing AI infrastructure environment. Located in San Francisco, this role offers competitive compensation and visa sponsorship.