Stand out for this role — generate a tailored resume and cover letter in about a minute.
salve.lab is seeking a Senior Site Reliability Engineer to own production reliability for services running on AWS and Amazon EKS. This fully remote role requires hands-on expertise in Kubernetes operations, incident management, and continuous improvement through automation and observability.
You will operate and troubleshoot production clusters, participate in on-call rotations, and drive improvements in SLIs, SLOs, and disaster recovery while communicating clearly with customers and teams.
salve.lab is seeking a Senior Site Reliability Engineer to own production reliability for services running on AWS and Amazon EKS. This fully remote role requires hands-on expertise in Kubernetes operations, incident management, and continuous improvement through automation and observability.
You will operate and troubleshoot production clusters, participate in on-call rotations, and drive improvements in SLIs, SLOs, and disaster recovery while communicating clearly with customers and teams.