An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Lewis Personnel Management seeks a Site Reliability Engineer (SRE) to maintain and evolve production infrastructure across AWS and Kubernetes. The role emphasizes reliability, scalability, automation, and incident management, with responsibilities for ML infrastructure.
You will operate a 24/7 Kubernetes-based production stack, lead incident responses, implement IaC with Terraform, Helm, and ArgoCD, and collaborate with data scientists to operationalize ML models at scale.
Lewis Personnel Management seeks a Site Reliability Engineer (SRE) to maintain and evolve production infrastructure across AWS and Kubernetes. The role emphasizes reliability, scalability, automation, and incident management, with responsibilities for ML infrastructure.
You will operate a 24/7 Kubernetes-based production stack, lead incident responses, implement IaC with Terraform, Helm, and ArgoCD, and collaborate with data scientists to operationalize ML models at scale.