Stand out for this role — generate a tailored resume and cover letter in about a minute.
Quix is seeking a Site Reliability Engineer to improve reliability and observability across enterprise platforms. You’ll define SLIs/SLOs, build observability infrastructure, and lead post-incident analyses.
This role emphasizes automation, capacity planning, and collaboration with engineering to ensure production readiness and reduce toil. Ideal candidates bring deep experience in distributed systems, on-call discipline, and strong scripting in Python, Bash, or Go.
Maintain and improve the reliability, observability, and operational maturity of enterprise platforms through systematic engineering discipline.
Quix is looking for a Site Reliability Engineer who approaches operational problems with an engineering mindset — building systems, tooling, and processes that prevent incidents rather than only responding to them. This role is for someone with deep experience in observability, incident management, capacity planning, and reliability engineering in distributed, production environments.
Work is calm, technical, and delivery-focused. You’ll help teams make durable decisions in enterprise environments where reliability and operational clarity matter.
Platform reliability is not an afterthought — it is a core commitment to clients who depend on Quix systems for operational continuity. The SRE function protects that commitment by building systems that fail gracefully, recover quickly, and improve systematically, reducing operational risk across every deployment.