Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Prime Intellect seeks a senior engineer to advance its open frontier AI infrastructure and developer platform. You will build distributed sandbox infra powering training systems and create intuitive web interfaces for AI workload management.
Key tech includes Rust, Linux, Prometheus/Grafana, TypeScript, React/Next.js, and RESTful APIs. Hybrid role with SF office preferred, visa sponsorship and relocation support offered.
Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team.
Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training, and deployment into one full-stack system for post-training at frontier scale - from SFT and RL to tool use, agent workflows, and continuously improving production models. We are building open frontier AI: open-source models trained end to end for long-horizon tasks like autonomous research, and the full-stack platform our own research team uses to build them. The next generation of AI companies, enterprises, and research teams do not just need more GPUs. They need the ability to turn their own workflows, tools, data, and feedback loops into superintelligence they own.
This is a hybrid role spanning both our infrastructure layers and developer platform. You'll work on two key areas:
The underlying sandbox infrastructure that powers our training systems
You will work on a distributed system with performance engineering at its core. The role will draw on the full breadth of your systems skills, from deep Linux kernel topics to high-level distributed system design. Expect your low-level systems fortitude to be pushed as you build infrastructure that remains fast, robust, and reliable at scale.
Infrastructure Development
Design and implement distributed orchestration infrastructure in Go and Rust
Build high-performance networking and coordination components
Create infrastructure automation pipelines with Ansible
Manage cloud resources and container orchestration
Implement scheduling systems for heterogeneous hardware (CPU, GPU, TPU)
Platform Development
Build intuitive web interfaces for AI workload management and monitoring
Develop REST APIs and backend services in Python
Create real-time monitoring and debugging tools
Implement user-facing features for resource management and job control
Infrastructure Skills
Systems programming experience with Rust
Strong Linux systems knowledge, including networking, namespacing, and performance tuning
Virtualization experience, including VMs, hypervisors, and low-level resource management
Observability tools (Prometheus, Grafana)
Platform Skills
Modern frontend development (TypeScript, React/Next.js, Tailwind)
Experience building developer tools and dashboards
RESTful API design and implementation
Experience with GPU computing and ML infrastructure
Knowledge of AI/ML model architecture and training
High-performance networking implementation
Open-source infrastructure contributions
Cash Compensation Range of $150-300k with significant equity incentives
Flexible work arrangement (San Francisco office preferred, remote possible for exceptional candidates)
Full visa sponsorship and relocation support
Professional development budget for courses and conferences
Regular team off-sites and conference attendance
Opportunity to shape the future of decentralized AI development
You’ll join a team of experienced engineers and researchers working on cutting‑edge problems in AI infrastructure. We believe in open development and encourage team members to contribute to the broader AI community through research and open-source contributions.
We value potential over perfection - if you’re passionate about democratizing AI development and have experience in either platform or infrastructure development (ideally both), we want to talk to you.