DevOps / Site Reliability Engineer (SRE) - London / Bracknell / Birmingham / Leeds / Manchester / Livingston - Hybrid - SCC Flex Contract
We are seeking an experienced DevOps / Site Reliability Engineer (SRE) to support the delivery, reliability and continuous improvement of critical technology platforms within a complex enterprise environment. This is an exciting opportunity for a skilled DevOps / SRE professional to join a high-profile client programme, driving automation, operational excellence and platform resilience across cloud and hybrid infrastructures.
The successful DevOps / Site Reliability Engineer (SRE) will play a key role in enabling scalable, secure and highly available services while promoting DevOps culture, reliability engineering principles and modern delivery practices.
Your responsibilities as the DevOps / Site Reliability Engineer (SRE):
- Design, implement and support highly available, scalable and resilient infrastructure platforms.
- Develop and maintain CI/CD pipelines to enable rapid, reliable and secure software delivery.
- Automate infrastructure provisioning, configuration management and deployment processes using Infrastructure as Code (IaC).
- Monitor platform health, availability, performance and capacity using modern observability and monitoring tools.
- Investigate, troubleshoot and resolve complex infrastructure and application issues.
- Implement reliability engineering practices to improve system stability, resilience and operational efficiency.
- Collaborate closely with development, engineering, architecture and operations teams to deliver robust platform solutions.
- Support cloud adoption and optimisation initiatives across Azure, AWS and hybrid environments.
- Drive continuous improvement through automation, standardisation and operational best practices.
- Participate in incident management, root cause analysis and post-incident review activities.
- Develop and maintain operational documentation, runbooks and support procedures.
- Ensure solutions align with organisational security, compliance, governance and operational standards.
- Contribute to platform roadmaps, technical strategy and transformation initiatives.
As a successful DevOps / Site Reliability Engineer (SRE), you will have:
- Proven experience working as a DevOps Engineer, Site Reliability Engineer (SRE) or similar platform engineering role within enterprise environments.
- Strong experience supporting cloud platforms including Microsoft Azure, AWS and/or Google Cloud Platform.
- Expertise in Infrastructure as Code tools such as Terraform, ARM, Bicep, CloudFormation or equivalent.
- Experience building and managing CI/CD pipelines using Azure DevOps, GitHub Actions, Jenkins or similar technologies.
- Strong experience with containerisation and orchestration technologies including Docker and Kubernetes.
- Knowledge of monitoring, observability and logging platforms such as Prometheus, Grafana, ELK, Dynatrace, Splunk or Datadog.
- Experience with scripting and automation using PowerShell, Python, Bash or similar languages.
- Strong understanding of Linux and Windows server administration.
- Knowledge of networking, security, identity management and enterprise infrastructure concepts.
- Experience implementing reliability engineering principles, service level objectives (SLOs) and service level indicators (SLIs).
- Strong analytical, troubleshooting and problem-solving skills.
- Excellent stakeholder engagement, collaboration and communication abilities.
- Experience working within Agile, DevOps and modern engineering practices.
- Relevant industry certifications in cloud, DevOps or platform engineering are highly desirable.
NOTE: At SCC, we take the privacy and security of your information very seriously. Any information we hold will be handled in accordance with current data protection legislation. Upon submitting your application, SCC will process your information in line with our privacy policy, which can be found on our website under Legal Privacy Notice Flexible Resourcing.