An application made for this job — a tailored resume and cover letter that speak straight to the posting.
CoreWeave seeks a Staff Software Engineer, Cluster Orch (Non SUNK), to lead architecture and technical vision for Kubernetes-native orchestration on massive GPU clusters. You will shape how AI workloads are admitted, scheduled, and governed at scale, while mentoring engineers and driving reliability across teams.
You will work with Go and distributed-system principles, contributing to a high-impact platform that underpins both training and inference workloads across CoreWeave’s global GPU
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com.
We're proud to be a Living Wage accredited Employer.
CoreWeave’s AI Workload Orchestration Platform team builds and operates the core, Kubernetes-native substrate that governs how massive AI workloads are admitted, scheduled, and executed across our global GPU footprint. Serving as a strategic complement to SUNK (Slurm on Kubernetes), this platform underpins both training and inference pipelines across the CoreWeave cloud, ensuring highly efficient resource utilisation for the world's most demanding AI applications.
As a Staff Software Engineer, Cluster Orch (Non SUNK), you will act as a principal technical leader driving CoreWeave’s Kubernetes-native orchestration strategy. You will own the technical vision and architecture for major portions of the platform, defining how AI workloads are admitted, scheduled, and governed across large GPU clusters using frameworks such as Kueue, Volcano, and Ray. In this high-impact role, you will apply systems thinking to resolve systemic performance, scalability, and fairness issues at scale. Additionally, you will lead cross-team architecture reviews, drive technical alignment across broader infrastructure, CKS, and managed inference teams, and establish platform-wide standards for reliability, capacity management, and developer experience while mentoring senior engineers across the organisation.
At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: