A complete application in a minute — tailored resume and cover letter, ready to send.
Salve.Lab is seeking a Senior Site Reliability Engineer to own production reliability on AWS and Amazon EKS. This is a hands-on role focused on operating highly available services, participating in on-call rotations, troubleshooting complex Kubernetes issues, and driving permanent improvements after incidents.
You will work with customers during escalations, define SLIs/SLOs, build runbooks, and implement IaC with Terraform/Terragrunt and GitOps tooling (Argo CD or FluxCD).
Salve.Lab is seeking a Senior Site Reliability Engineer to own production reliability on AWS and Amazon EKS. This is a hands-on role focused on operating highly available services, participating in on-call rotations, troubleshooting complex Kubernetes issues, and driving permanent improvements after incidents.
You will work with customers during escalations, define SLIs/SLOs, build runbooks, and implement IaC with Terraform/Terragrunt and GitOps tooling (Argo CD or FluxCD).