Get more replies from employers
Send a job-specific resume in minutes.
Lambda in Bellevue/ SF/ SJ is seeking a Staff Software Engineer to define the technical vision for the next-gen GPU/CPU host lifecycle and compute control plane. You will lead cross-functional teams, bridging BIOS/firmware, Linux kernel internals, DPU utilization, and cloud provisioning at massive scale.
You will guide crucial architectural decisions, ensure reliability, and establish standards for enterprise-grade cloud software, while mentoring senior engineers across multiple teams.
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.
If you'd like to build the world's best AI cloud, join us.
*Note: This position requires presence in our Bellevue, San Francisco, or San Jose office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.
As a Staff Software Engineer for the Compute pillar, you will play a critical role in defining the technical vision for Lambda's next-generation GPU and CPU host instance lifecycle and compute control plane. This role bridges the gap between high-level distributed systems and low-level semiconductor architecture to enable seamless, reliable cloud provisioning and lifecycle management of a heterogeneous compute platform at a massive scale. You will provide hands‑on technical leadership that will guide development of a resilient compute control plane utilizing durable execution concepts and deep/unique hardware integration.The position requires a deep understanding of the entire stack, from BIOS/firmware (UEFI), Linux kernel internals, modern DPU capabilities, distributed systems, cradle-to-grave system lifecycle management, to large-scale cloud-service provider (CSP) operations. You will drive high‑impact, cross‑functional initiatives, leading the work of multiple engineers to deliver enterprise‑grade SLAs for the world's leading AI researchers.
We are seeking an engineer with extensive experience in cloud infrastructure to build and optimize GPU-first compute systems. In this role, you will be responsible for:
The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.
Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.