We're working with a growing technology business that's transforming the way organisations manage critical data through a highly scalable cloud platform.
As they continue to invest in their infrastructure and engineering capability, they're looking for a Senior Site Reliability Engineer to help build, maintain, and optimise the environments that support their mission-critical applications.
The Why? (A Couple of Highlights)
- Work with modern cloud technologies – Join a team focused on automation, scalability, and reliability across a cloud-first platform.
- Make a real impact – Help maintain and improve a platform that supports customers globally, ensuring high availability and performance.
The Role
You'll be responsible for supporting and improving the environments that power a large-scale SaaS platform, working closely with software engineers and technical teams to deliver reliable, secure, and scalable infrastructure.
This is a hands-on role where you'll help drive automation, improve system resilience, and contribute to the ongoing evolution of the platform.
What You'll Be Doing
- Building, maintaining, and supporting cloud-hosted environments.
- Monitoring platform health and implementing alerting and self-healing solutions.
- Managing Kubernetes environments and containerised workloads.
- Troubleshooting infrastructure and application issues alongside development teams.
- Designing scalable infrastructure solutions to support future growth.
- Maintaining technical documentation and architectural diagrams.
- Mentoring junior engineers and promoting engineering best practices.
- Driving continuous improvements across infrastructure and deployment processes.
What We're Looking For
- Strong experience with Kubernetes and container technologies.
- Experience administering both Linux and Windows environments.
- Commercial experience with AWS, Azure, GCP, or Oracle Cloud.
- Strong scripting skills using PowerShell and/or Bash.
- Experience with CI/CD and configuration management tools such as Jenkins, Helm, or Ansible.
- Knowledge of relational databases including Oracle or SQL Server.
- A proactive engineer who enjoys automation, problem-solving, and continuously improving platform reliability.
- Experience supporting .NET applications.
- Exposure to AI/ML tools used for automation, observability, or incident management.
What's in It for You?
- Comprehensive healthcare and benefits package.
- Flexible working environment with an emphasis on work-life balance.
- The opportunity to work on a large-scale cloud platform using modern technologies.
- Join a collaborative engineering team where you'll have genuine influence over the future of the platform.