An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Evlo AI is seeking a Site Reliability Engineer to own the reliability, scalability, and operational readiness of production systems on AWS and Kubernetes. You will automate infrastructure, improve observability, and lead incident response to minimize downtime and downtime impact.
You will collaborate with platform, application, and security engineers to improve service availability, deployment safety, and recovery time.
The Site Reliability Engineer owns the reliability, scalability, and operational readiness of production systems running across AWS and Kubernetes. The role focuses on infrastructure automation, observability, incident response, and eliminating recurring operational work through engineering.
The Site Reliability Engineer owns the reliability, scalability, and operational readiness of production systems running across AWS and Kubernetes. The role focuses on infrastructure automation, observability, incident response, and eliminating recurring operational work through engineering.
You will partner with platform, application, and security engineers to improve service availability, deployment safety, and recovery time. The team manages systems where measurable uptime, predictable performance, and disciplined production operations are critical to customer experience.