Site Reliability Engineer - Cloud, Kubernetes & Observability
Insight International (UK) Ltd
Bournemouth
On-site
GBP 55,000 - 75,000
Full time
14 days+
Application generator
A complete application in a minute — tailored resume and cover letter, ready to send.
Get past ATS filters
Job summary
A leading technology firm based in Bournemouth is looking for a Site Reliability Engineer to design and maintain highly available Java/Spring Boot applications. You will focus on cloud migration, automating infrastructure using Kubernetes and Terraform, and driving analytics for system performance optimization. Ideal candidates will have experience in large-scale systems and strong collaboration skills. This permanent role is onsite five days a week.
Qualifications
Proven experience as a Site Reliability Engineer, DevOps Engineer, or similar role supporting large‑scale, mission‑critical systems.
Strong hands‑on experience with Kubernetes and Terraform.
Experience deploying and operating applications in public cloud environments (AWS, Azure, GCP).
Solid understanding of Java and Spring Boot applications.
Experience with monitoring, logging, and observability tools (Prometheus, Grafana, ELK, Splunk).
Strong troubleshooting and problem‑solving skills.
Excellent communication and collaboration skills.
Responsibilities
Design, implement, and maintain systems that are robust, scalable, and highly available.
Lead and support migration of applications and infrastructure to public cloud platforms.
Develop and maintain automation scripts and infrastructure using Kubernetes and Terraform.
Build and enhance monitoring, alerting, and observability solutions.
Partner with software engineers, product managers, and business stakeholders to deliver solutions.
Leverage cloud‑based analytics tools to monitor system health and optimize performance.
Identify and implement opportunities to improve reliability, efficiency, and scalability.
Skills
Site Reliability Engineering
Kubernetes
Terraform
Java
Spring Boot
Cloud environments
Monitoring tools
Troubleshooting
Collaboration
Tools
AWS
Azure
GCP
Prometheus
Grafana
ELK
Splunk
Jenkins
GitHub Actions
Job description
A leading technology firm based in Bournemouth is looking for a Site Reliability Engineer to design and maintain highly available Java/Spring Boot applications. You will focus on cloud migration, automating infrastructure using Kubernetes and Terraform, and driving analytics for system performance optimization. Ideal candidates will have experience in large-scale systems and strong collaboration skills. This permanent role is onsite five days a week.