An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Piris Labs is building the next generation of AI inference infrastructure. We are seeking a Founding AI Infrastructure Engineer to architect and build our inference platform from first principles.
This role spans the entire stack—from control plane orchestration to distributed serving systems, networking, scheduling, observability, and performance optimization. As an early engineer, you will have enormous influence over our architecture, engineering culture, hiring strategy, and product
We are building the next generation of AI inference infrastructure. Our mission is to make large-scale model serving dramatically faster, more efficient, and more reliable by tightly integrating software orchestration with next-generation GPU and networking architectures. We believe the future of AI infrastructure will be defined not just by larger models, but by how effectively those models are distributed across nodes, networks, regions, and heterogeneous hardware. We're assembling a small, high-caliber engineering team to build that future from the ground up.
We are looking for a Founding AI Infrastructure Engineer to help architect and build our inference platform from first principles. This is not a traditional backend role. You will work across the entire inference stack, from control plane orchestration to distributed serving systems, networking, scheduling, observability, and performance optimization. You will help solve some of the hardest problems in AI infrastructure. As one of the first engineers, you will have enormous influence over our architecture, engineering culture, hiring strategy, and product direction. This role is designed for builders who want to create foundational technology, not maintain existing systems.
Strong Technical Foundations: exceptional fundamentals in distributed systems, networking, operating systems, databases, concurrency, and systems design. Builder Mentality: you enjoy building new systems from scratch and taking ownership of ambiguous, high-impact problems. Programming Excellence: strong experience in one or more of Go, Rust, C++, or Python. Infrastructure Experience: experience with Kubernetes, Linux systems, containers, service meshes, cloud infrastructure, and infrastructure automation. High Learning Velocity: we care more about your ability to learn and execute than your years of experience.
AI infrastructure companies, distributed databases, high-performance networking, cloud infrastructure platforms, GPU systems, HPC environments, Kubernetes platforms, and large-scale backend systems. Experience with technologies such as Kubernetes, Envoy, eBPF, OpenTelemetry, vLLM, SGLang, TensorRT-LLM, NCCL, RDMA, and InfiniBand is a strong plus, but not required.
Within your first year, you will help build a production-grade distributed inference platform, multi-node model serving capabilities, intelligent global request routing, automated GPU fleet management, and the foundation of our engineering organization.
Compensation: $100,000–$200,000 depending on experience, plus equity, 401(k), health insurance, and benefits keyvan@pirislabs.io