A complete application in a minute — tailored resume and cover letter, ready to send.
AlloFresh is seeking a Site Reliability Engineer to ensure reliability, scalability, and performance of production systems in a cloud-native environment. You will implement SRE practices, automate infrastructure, and lead incident response while partnering with development teams to embed reliability across the software lifecycle.
You will manage Kubernetes clusters, maintain GitOps workflows with ArgoCD, design CI/CD pipelines with Jenkins, and contribute to runbooks, postmortems, and
As a Site Reliability Engineer, you will be responsible for ensuring the reliability, scalability, and performance of production systems. You will apply engineering principles to operations, focusing on automation, observability, and continuous improvement to reduce manual effort and enhance system resilience. You will work closely with development teams to build and operate highly reliable services, embedding reliability into the full software lifecycle.