Get more replies from employers
Send a job-specific resume in minutes.
Lambda seeks a Senior Site Reliability Engineer to scale and harden its AI cloud platform across data centers. You’ll improve provisioning reliability, implement robust monitoring, and drive incident response and postmortems.
You will define SLIs/SLOs, automate drift remediation, and build disaster recovery workflows using Terraform, Argo CD, and Kubernetes-native tools. Collaboration with multiple teams and mentorship are key aspects of the role.
Lambda seeks a Senior Site Reliability Engineer to scale and harden its AI cloud platform across data centers. You’ll improve provisioning reliability, implement robust monitoring, and drive incident response and postmortems.
You will define SLIs/SLOs, automate drift remediation, and build disaster recovery workflows using Terraform, Argo CD, and Kubernetes-native tools. Collaboration with multiple teams and mentorship are key aspects of the role.