An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Prime Intellect is building the open superintelligence stack, delivering a full-stack platform for post-training at frontier scale. You will design GPU cluster architectures, plan capacity for large-scale deployments, and develop deployment strategies for LLM training and HPC workloads.
You will optimize networking, file systems, and driver stacks while maintaining 24/7 support for critical customer deployments.
Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team.
Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training, and deployment into one full-stack system for post-training at frontier scale - from SFT and RL to tool use, agent workflows, and continuously improving production models. We are building open frontier AI: open-source models trained end to end for long-horizon tasks like autonomous research, and the full-stack platform our own research team uses to build them. The next generation of AI companies, enterprises, and research teams do not just need more GPUs. They need the ability to turn their own workflows, tools, data, and feedback loops into superintelligence they own.
You'll work directly with customers pushing the boundaries of AI, from startups training foundation models to enterprises deploying massive inference infrastructure. You'll collaborate with our world-class engineering team while having direct impact on systems powering the next generation of AI breakthroughs.
We value expertise and customer obsession - if you're passionate about building reliable, high-performance GPU infrastructure and have a track record of successful large-scale deployments, we want to talk to you.
Cash Compensation Range of $150-300k plus Equity Incentives