Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Ardan Labs is seeking a Senior Site Reliability Engineer to join a platform team and take ownership of reliability for an internal developer platform and for the experiences of application teams building on it.
This hands-on senior individual contributor role focuses on operating Kubernetes at scale, solving complex infrastructure problems, and working directly with teams to diagnose and resolve issues, including HIPAA-compliant security hygiene and patching as ongoing priorities.
We are looking for a Senior Site Reliability Engineer to join a platform team and take ownership of both the reliability of an internal developer platform and the experience of the application teams building on it.
This is a hands-on senior individual contributor role for an engineer who enjoys solving complex infrastructure problems, operating Kubernetes at scale, and working directly with teams to diagnose and resolve issues.
You will work across Amazon EKS and Red Hat OpenShift on AWS, a curated Helm chart catalog, GitOps pipelines, Terraform, and a multi-account AWS environment built on Control Tower. Our workloads operate under HIPAA requirements, making security hygiene, patch currency, and operational reliability ongoing engineering priorities.
Approximately 60% of your time will focus on reliability engineering and 40% on forward-deployed work with application, security, network, and other technical teams.
This isn't a role where success is measured by how many tools you know. The platform is highly automated and largely self-healing. The difficult problems are often about figuring out where the problem actually belongs, determining the right durable fix, and building agreement between teams that don't report to one another.
We're looking for an engineer who can operate at both levels: deep technical infrastructure expertise and strong cross-team engagement.