Senior Site Reliability Engineer

Tenth Revolution Group

Knutsford

Hybrid

GBP 70,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading Financial Services firm is seeking a Senior Site Reliability Engineer for its Knutsford office. In this hybrid role, you will enhance SRE best practices, ensuring system availability and performance. Key responsibilities include automating processes, responding to incidents, and collaborating with teams to improve reliability. Ideal candidates will have expertise in programming, incident management, and cloud computing.

Responsibilities

  • Ensure availability and performance of systems through monitoring and maintenance.
  • Respond to outages and implement preventive measures.
  • Develop tools to automate operational processes.
  • Monitor system performance and resource usage.
  • Collaborate with development teams for best practices integration.

Skills

Programming and Scripting (Python, Powershell, Go)
Incident Management and Troubleshooting
Systems Engineering and Automation
Communication Skills
Cloud Computing Knowledge

Job description

Overview

Senior Site Reliability Engineer — Knutsford (hybrid, 2 days per week in office). A leading Financial Services firm is recruiting for a Senior Site Reliability Engineer to become part of a newly formed Core SRE Team that will establish a Centre of Excellence to enhance and promote SRE best practices.

About the Role

As a key hire, you will raise awareness and drive adoption of SRE methodologies within various teams. This is a hands-on engineering role where you will design, build, and optimise automation frameworks, observability tools, and incident response mechanisms. You will act as a trusted advisor, providing strategic guidance and consultative support to help teams improve reliability, scalability, and efficiency.

Responsibilities

  • Availability, performance, and scalability of systems and services through proactive monitoring, maintenance, and capacity planning.
  • Resolution, analysis and response to system outages and disruptions, and implementation of measures to prevent similar incidents from recurring.
  • Development of tools and scripts to automate operational processes, reducing manual workload, increasing efficiency, and improving system resilience.
  • Monitoring and optimisation of system performance and resource usage, identifying bottlenecks, and implementing best practices for performance tuning.
  • Collaboration with development teams to integrate best practices for reliability, scalability, and performance into the software development lifecycle, and work closely with other teams to ensure smooth and efficient operations.

Required Skills

  • Proficiency in Programming and Scripting — languages such as Python, Powershell, or Go for automating routine tasks and system deployments.
  • Incident Management and Troubleshooting — ability to manage incidents effectively, troubleshoot issues swiftly, and perform root cause analysis to prevent future incidents.
  • Systems Engineering and Automation — understanding of operating systems, networking, and cloud infrastructure; proficiency in automation tools for maintaining system reliability at scale.
  • Influential Communication Skills — ability to communicate effectively with team members and stakeholders to drive alignment and foster a collaborative environment for SRE practices.
  • Knowledge of Cloud Computing — familiarity with cloud platforms and services as infrastructure moves to the cloud.

Seniority level

  • Mid-Senior level

Employment type

  • Full-time

Job function

  • Information Technology

Industries

  • Technology, Information and Media
  • Financial Services
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Xpertise Recruitment • West Drayton

On-site
GBP 60,000 - 80,000
Site Reliability Engineer
Site Reliability Engineer

DNS INFO LTD • City Of London

On-site
GBP 70,000 - 95,000
Senior Site Reliability Engineer - Drive Automation & Reliability
Senior Site Reliability Engineer - Drive Automation & Reliability

Tenth Revolution Group • Knutsford

Hybrid
GBP 70,000 - 90,000
Site Reliability Engineer
Site Reliability Engineer

ScaleneWorks People Solutions LLP • Bournemouth

On-site
GBP 60,000 - 80,000
Senior SRE
Senior SRE

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LSEG • Nottingham

On-site
GBP 70,000 - 90,000
Healthcare
Retirement planning
Paid volunteering days
+1
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Reward Gateway • Greater London

Hybrid
GBP 60,000 - 65,000
Hybrid work option
Site Reliability Engineer
Site Reliability Engineer

Computappoint • City Of London

Hybrid
GBP 56,000 - 75,000
SRE Technical Lead
SRE Technical Lead

83zero Ltd • Wokingham

Hybrid
GBP 60,000 - 100,000
5% bonus
Hybrid working model