Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Acceler8 Talent in San Francisco is seeking a Member of Technical Staff, Infrastructure to build and operate the cluster infrastructure behind our AI platform. You will determine how new accelerator hardware is brought online, how compute fleets are provisioned, and how production inference systems remain reliable at scale.
You will deploy production clusters across accelerator architectures, automate provisioning and fleet lifecycle, improve scheduling and resource utilization, and build
I’m working with a well-funded AI infrastructure startup (Series A) building a cloud platform that runs inference workloads across GPUs, CPUs, and emerging accelerator architectures.
The team is solving a difficult infrastructure problem: making new different hardware accelerators usable through one reliable platform, without requiring customers to redesign their software stack for every accelerator.
This role will build the cluster infrastructure behind that platform. You’ll determine how new hardware is brought online, how compute fleets are provisioned and operated, and how production inference systems remain reliable as they scale.
This is an opportunity to join a small, highly technical team and build production infrastructure across multiple generations and types of AI hardware. The strongest candidates will be able to explain what they personally built, operated, measured, and debugged at scale.