Stand out for this role — generate a tailored resume and cover letter in about a minute.
Evlo AI is seeking a Site Reliability Engineer to own the reliability, scalability, and performance of production distributed systems across multi-region cloud environments. You will work with software squads to embed resilience, automate toil, and maintain strict SLAs.
You will design and maintain infrastructure on AWS or GCP using Terraform and Pulumi, implement observability with Prometheus, Grafana, OpenTelemetry, and Datadog, and drive incident response with blameless post-mortems.
Evlo AI is seeking a Site Reliability Engineer to own the reliability, scalability, and performance of production distributed systems across multi-region cloud environments. You will work with software squads to embed resilience, automate toil, and maintain strict SLAs.
You will design and maintain infrastructure on AWS or GCP using Terraform and Pulumi, implement observability with Prometheus, Grafana, OpenTelemetry, and Datadog, and drive incident response with blameless post-mortems.