Lead Site Reliability Engineer - Glasgow

Hackajob Ltd

Glasgow

On-site

GBP 90,000 - 110,000

Full time

36 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hackajob Ltd. is seeking an experienced Site Reliability Engineer in the UK (Scotland/Glasgow) to champion reliability across teams and architectures. You will influence incident response, observability practices, and platform resilience while mentoring engineers and driving standards.

The role emphasizes collaboration with stakeholders, SLIs/SLOs, and a strong bias for action in large-scale enterprise environments. This position offers a competitive salary and on-site work in Glasgow.

Qualifications

  • Deep proficiency in reliability, scalability, and security practices.
  • Fluency in Python/Java/Spring Boot/.NET or similar languages.
  • Hands-on observability with monitoring and telemetry tools.
  • Experience with CI/CD tools (Jenkins, GitLab, Terraform).

Responsibilities

  • Demonstrate and champion site reliability culture and practices, and exert technical influence throughout our team.
  • Lead initiatives to improve the reliability and stability of our team's applications and platforms using data-driven analytics to improve service levels.
  • Collaborate with team members to identify SLIs/SLOs and establish reasonable error budgets with customers.
  • Demonstrate a high level of technical expertise and proactively solve bottlenecks in our domains.
  • Act as the main contact during major incidents and resolve issues quickly to avoid losses.
  • Document and share knowledge within our organization via internal forums and communities of practice.
  • Take the lead on resiliency design reviews.
  • Break up complex problems into actionable work for engineers.
  • Act as a technical lead for medium to large-sized products.
  • Provide advice and mentoring to other engineers.

Skills

SRE fundamentals
Observability
CI/CD
Containerization
Networking
Mentoring
Cloud platforms

Tools

Grafana
Dynatrace
Prometheus
Datadog
Splunk
Jenkins
GitLab
Docker
Kubernetes
Terraform

Job description

Salary: £100,000 - 100,000 per year

Requirements
  • Deep proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practices, with the ability to implement these practices within an application or platform
  • Fluency in at least one programming language, such as Python, Java Spring Boot, or .NET
  • Deep knowledge of software applications and technical processes with emerging depth in one or more technical disciplines
  • Proficiency and hands-on experience in observability practices, including white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, or Splunk
  • Proficiency in continuous integration and continuous delivery tools such as Jenkins, GitLab, or Terraform
  • Experience with container technologies and container orchestration platforms such as ECS, Kubernetes, or Docker
  • Experience troubleshooting common networking technologies and issues
  • Ability to identify and resolve problems related to complex data structures and algorithms
  • Ability to collaborate and communicate effectively across different levels and stakeholder groups
  • Experience mentoring or coaching engineers on site reliability practices and engineering standards
  • Familiarity with cloud platforms and infrastructure-as-code practices in large-scale enterprise environments
  • Experience contributing to or leading communities of practice, internal knowledge sharing, or engineering guilds
  • Exposure to chaos engineering or fault injection methodologies to proactively test system resilience
  • Ability to evaluate and introduce emerging technologies that improve platform reliability and reduce operational toil
Responsibilities
  • Demonstrate and champion site reliability culture and practices, and exert technical influence throughout our team
  • Lead initiatives to improve the reliability and stability of our teams applications and platforms using data-driven analytics to improve service levels
  • Collaborate with team members to identify comprehensive service level indicators and establish reasonable service level objectives and error budgets with customers
  • Demonstrate a high level of technical expertise within one or more technical domains and proactively identify and solve technology-related bottlenecks in our areas of expertise
  • Act as the main point of contact during major incidents for our application and identify and solve issues quickly to avoid financial losses
  • Document and share knowledge within our organization via internal forums and communities of practice
  • Take the lead on resiliency design reviews
  • Break up complex problems into digestible work for other engineers
  • Act as a technical lead for medium to large-sized products
  • Provide advice and mentoring to other engineers
Technologies
  • Cloud
  • Datadog
  • Docker
  • Dynatrace
  • GitLab
  • Grafana
  • Java
  • Jenkins
  • Kubernetes
  • Marketing
  • Prometheus
  • Python
  • Security
  • Splunk
  • Spring
  • Spring Boot
  • Terraform
  • ASP.NET
  • DevOps

last updated 36 week of 2026

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer - Edinburgh
Lead Site Reliability Engineer - Edinburgh

Inspire People • City of Edinburgh

Hybrid
GBP 72,000 - 88,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VIQU IT Recruitment • Kingston

On-site
GBP 68,000 - 83,000
On-call allowance
Bonus
Lead Site Reliability Engineer - Glasgow
Lead Site Reliability Engineer - Glasgow

JP Morgan Chase • Glasgow

On-site
GBP 62,000 - 102,000
Site Reliability Engineer (SRE) – Cloud Engineer
Site Reliability Engineer (SRE) – Cloud Engineer

Dianaduggan • Glasgow

Hybrid
GBP 74,000 - 83,000
Lead SRE - AWS Platform
Lead SRE - AWS Platform

JP Morgan Chase • Glasgow

On-site
GBP 62,000 - 102,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Lead Site Reliability Engineer: Architect & Incident Leader
Lead Site Reliability Engineer: Architect & Incident Leader

Hackajob Ltd • Glasgow

On-site
GBP 90,000 - 110,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

McLaren Automotive Ltd • Woking

On-site
GBP 70,000 - 90,000
25 days holiday plus bank holidays
Enhanced company pension scheme
Discretionary annual bonus
+6
Senior Lead Site Reliability / DevOps Engineer
Senior Lead Site Reliability / DevOps Engineer

JP Morgan Chase • Glasgow

On-site
GBP 62,000 - 102,000
Site Reliability Engineer
Site Reliability Engineer

ScaleneWorks People Solutions LLP • Bournemouth

On-site
GBP 60,000 - 80,000