Our client is looking an experienced Site Reliability Engineer (SRE) to join our team and help ensure the reliability, scalability, performance and operational resilience of enterprise security platforms.
Key Responsibilities
- Drive SRE and reliability engineering practices across cyber security platforms, improving availability, resilience, scalability and performance.
- Automate operational processes and eliminate toil through engineering solutions, Infrastructure as Code and automation-first approaches.
- Develop observability, monitoring, alerting and reliability metrics, including SLIs, SLOs and actionable operational dashboards.
- Lead platform reliability and continuous improvement initiatives, identifying risks and improving operational readiness and service stability.
- Support incident analysis and post-incident reviews, driving preventative improvements and greater platform resilience.
- Partner with cyber security, platform engineering and technology teams, influencing engineering standards, operating models and long-term platform maturity.
Our client is looking an experienced Site Reliability Engineer (SRE) to join our team and help ensure the reliability, scalability, performance and operational resilience of enterprise security platforms.
Key Responsibilities
- Drive SRE and reliability engineering practices across cyber security platforms, improving availability, resilience, scalability and performance.
- Automate operational processes and eliminate toil through engineering solutions, Infrastructure as Code and automation-first approaches.
- Develop observability, monitoring, alerting and reliability metrics, including SLIs, SLOs and actionable operational dashboards.
- Lead platform reliability and continuous improvement initiatives, identifying risks and improving operational readiness and service stability.
- Support incident analysis and post-incident reviews, driving preventative improvements and greater platform resilience.
- Partner with cyber security, platform engineering and technology teams, influencing engineering standards, operating models and long-term platform maturity.
Skills & Experience Required
- 5+ years’ experience in SRE, DevOps or Platform Engineering within a large enterprise environment.
- A Bachelor’s or master’s degree in computer science, Software Engineering, Information Technology or a related technical discipline, or equivalent practical experience.
- Demonstrated experience implementing SRE principles, particularly within security operations or security platform engineering.
- Strong hands-on experience with SLIs, SLOs, observability, incident management and post-incident reviews.
- Strong Python scripting and automation capabilities.
- Experience with Terraform or equivalent Infrastructure as Code tooling.
- Experience with CI/CD technologies and operational configuration management.
- Cloud engineering experience across at least one major cloud provider: AWS, Azure or GCP.
- Palo Alto Cortex XSIAM, Microsoft Defender for Endpoint, Tenable or equivalent enterprise security platforms preferred but not essential
Peoplebank and Leaders IT are committed to creating a diverse and inclusive workplace where everyone belongs. We welcome applications from people of all backgrounds, identities, and experiences. If you need adjustments to the recruitment process due to your circumstances, please let us know—we’re here to support you.