Get more replies from employers
Send a job-specific resume in minutes.
Lambda, The Superintelligence Cloud, seeks an experienced Senior Software Engineer to join the Storage team and help build next-generation on-premise storage software for AI workloads. You will shape storage software architecture, APIs, and resilience across file, block, and object storage.
You will collaborate with hardware teams to integrate NVMe and GPU-direct storage, tackle complex distributed systems challenges, and drive performance, reliability, and scalability across large data sets.
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.If you'd like to build the world's best AI cloud, join us.
This position requires presence in our San Francisco, San Jose, or Bellevue office location 4 days per week; Lambda’s designated work-from-home day is currently Tuesday.
In the world of distributed AI training and inference, raw GPU and CPU horsepower is just a part of the story. High-performance networking and storage are the critical components that enable and unite these systems, making groundbreaking AI training and inference possible.
The Lambda Infrastructure Engineering organization forges the foundation of high-performance AI clusters by welding together the latest in AI storage, networking, GPU and CPU hardware.
AI training and inference relies on petabytes of data hosted on large, high-performance storage arrays. At Lambda, the Infrastructure Storage Team’s job is to ensure that the data powering AI is fast, performant, and available across a variety of access protocols (fit for purpose).
We're looking for an experienced Senior Software Engineer to join our storage team. You'll join a team responsible for developing and implementing storage software for our next-generation on-premise storage solutions. This role requires expertise in distributed systems, and an in-depth understanding of file, block, and object storage protocols. You'll work on building scalable and resilient storage services that power our AI and machine learning infrastructure.