Get more replies from employers
Send a job-specific resume in minutes.
AMP PBC is building the largest independent AI fleet. You will own the east-west network across clusters, bringing up the fabric from switch access to production-grade performance for high-scale GPU training workloads.
This founding role requires hands-on deployment, tuning, and leadership to ensure robust, scalable network infrastructure. You will work with RoCEv2/RDMA at scale, diverse switch platforms, and NCCL benchmarking, shaping the fabric for frontier AI labs.
AMP PBC is the AI infrastructure partner to independent teams building at the frontier. AMP is a public benefit company comprised of two major divisions, which together deliver both compute and capital under management.
AMP PBC is the AI infrastructure partner to independent teams building at the frontier. AMP is a public benefit company comprised of two major divisions, which together deliver both compute and capital under management.
AMP's technology division, AMP Infra, is building the independent AI Grid: pooled, automated infrastructure orchestration, across clouds and other compute providers, to give frontier teams on-demand access to the highest quality compute at any scale. Having built internal solutions for the world's largest hyperscalers, the AMP team is now creating a global, silicon-agnostic infrastructure network so that any team has the compute resources to build at the frontier without giving up their independence.
AMP's venture arm, AMP Foundry, partners with the world's leading researchers and scientists, incubating ideas and deploying strategic capital into frontier labs and other key areas of the AI infrastructure stack. With over $1 billion under management, Foundry operates at the pace the frontier requires, and provides fuel to help the best teams push the scaling laws.
AMP is backed by world-class investors, and partnered with leading labs, hyperscalers, research institutions, chipmakers and compute providers. We offer deep expertise, genuine ownership, and a relentless drive to maximize the world's frontier output.
We are building the largest independent AI fleet in the world, and the fabric is what decides whether it works. You will own the east-west network across our clusters: the GPU-to-GPU interconnect, RoCEv2 and InfiniBand, from switch access on day one through a tuned cluster that trains at the performance our customers paid for. This is the founding role on our network team. The industry has undervalued how hard it is to bring up a cluster properly, and networking is the most undervalued part of it. A deployment can be enormous and enormously expensive, and if the fabric is wrong it was all for nothing. We treat this as a core competency of the business, not a support function, and it is top of mind for the founders.
If you're driven to build the infrastructure that lets the world's best teams push the frontier, you belong here.