Erhalte mehr Antworten von Arbeitgebern
Versende in nur wenigen Minuten einen passgenauen Lebenslauf.
Lambda, The Superintelligence Cloud, seeks a Senior Site Reliability Engineer to bolster reliability, scalability, and operational maturity of our cloud platform. You will work across Kubernetes, infrastructure automation, observability, deployment systems, and incident response to build a resilient foundation for AI workloads.
Join Lambda’s Core Cloud Platform to provision compute, orchestrate infrastructure, and scale services across data centers as the fleet grows, with a focus on reliability
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.
*Note: This position requires presence in our San Francisco, San Jose, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.
About the Role Lambda’s Core Cloud Platform powers compute provisioning and infrastructure orchestration across our physical data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability, scalability, and operational maturity of these systems as Lambda’s fleet and customer base grow. You will work across Kubernetes, infrastructure automation, observability, deployment systems, and incident response to build a resilient foundation for customer AI workloads.