SRE Operations Support

TEKsystems

Greater London

On-site

GBP 55,000 - 65,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TEKsystems in London, UK is seeking an experienced SRE Operations Support to act as a Subject Matter Expert within a Site Reliability Engineering team. You will manage, deploy, troubleshoot and monitor production systems to ensure reliability and performance across environments.

Working closely with support and engineering, you will implement automation, develop deployment guidelines, write SQL scripts, participate in incident management, and mentor teammates while maintaining documentation.

Qualifications

  • 5+ years' experience troubleshooting applications in production or operational environments.
  • 3+ years' experience writing and supporting SQL, with strong SQL development skills.
  • Proficient in Microsoft SQL Server and SSMS.
  • 3-5+ years' experience in Systems Engineering, including hands-on experience with Windows Server.
  • Experience working in Agile environments using JIRA and SCRUM methodologies.
  • Proven ability to troubleshoot issues across application and hosting environments, including multi-tier architectures.
  • Enthusiasm for learning new technologies and applying knowledge quickly.
  • Strong teamwork, communication and stakeholder management skills.
  • Good to advanced practical SQL programming experience.
  • Excellent problem-solving and analytical skills across systems and applications.
  • Working knowledge of Windows operating systems within enterprise environments.

Responsibilities

  • Deploy applications across all production environments, including Staging, Pre-Production, UAT and Production.
  • Troubleshoot application and system issues, identifying root causes and implementing effective solutions.
  • Perform testing and quality assurance activities following deployments to validate functionality and stability.
  • Develop and enhance testing and deployment guidelines to support consistent and reliable releases.
  • Monitor client systems to ensure services operate as expected and performance stays within thresholds.
  • Implement automation processes using proprietary worker and job management systems to improve efficiency.
  • Collaborate with support teams to triage, investigate and resolve issues impacting live systems.
  • Create SQL queries and scripts to process workloads and support operational requirements.
  • Investigate system performance issues and recommend optimisation opportunities.
  • Support day-to-day operation of systems through proactive monitoring and early issue detection.
  • Work with support teams on analysis, incident management and problem resolution.
  • Apply strong analytical and problem-solving skills when investigating issues.
  • Maintain a structured, disciplined approach to work and documentation.
  • Adhere to team processes and workflows, contributing to a collaborative environment.
  • Develop knowledge across systems and technologies, adapting quickly to new tools.
  • Communicate effectively with cross-functional teams when requesting or providing support.
  • Prioritise client operations and service reliability in daily activities and decisions.
  • Perform management responsibilities including interviewing, hiring and training employees.
  • Plan, allocate and oversee work for team members, support performance management and development.
  • Assist with resolving team challenges and fostering a positive, collaborative environment.
  • Carry out additional duties to support reliability, availability and performance of production systems.

Skills

SQL
SSMS
Windows Server
Agile
Troubleshooting
Automation
Monitoring
Incident Mgmt
Communication
Teamwork
C#
AWS

Education

Bachelor's degree in Engineering/Computer Science

Tools

JIRA
IIS
ECS
AWS

Job description

Job Description

The SRE Operations Support role acts as a Subject Matter Expert within a dedicated Site Reliability Engineering team, responsible for managing, maintaining and troubleshooting complex production environments. You will support multiple stages of the deployment lifecycle, ensuring applications are delivered reliably, thoroughly tested and effectively monitored. Working closely with both support and engineering teams, you will help maintain the stability, performance and reliability of mission-critical systems.


Responsibilities


  • Deploy applications across all production environments, including Staging, Pre-Production, UAT and Production.

  • Troubleshoot application and system issues, identifying root causes and implementing effective solutions.

  • Perform testing and quality assurance activities following deployments to validate functionality and stability.

  • Develop and enhance testing and deployment guidelines to support consistent and reliable releases.

  • Monitor client systems to ensure services are operating as expected and performance remains within agreed thresholds.

  • Implement automation processes using proprietary worker and job management systems to improve operational efficiency.

  • Collaborate with support teams to triage, investigate and resolve issues impacting live systems.

  • Create SQL queries and scripts to process workloads and support operational requirements.

  • Investigate system performance issues and recommend optimisation opportunities where appropriate.

  • Support the day-to-day operation of systems and products through proactive monitoring and early issue detection.

  • Work closely with support teams to assist with analysis, incident management and problem resolution.

  • Apply strong analytical and problem-solving skills when investigating system and application issues.

  • Maintain a structured and disciplined approach to work, ensuring tasks are completed and appropriately documented.

  • Adhere to established team processes and workflows, contributing to a consistent and collaborative way of working.

  • Develop knowledge across a range of systems and technologies, adapting quickly to new tools and platforms.

  • Communicate effectively with cross-functional teams when requesting or providing support.

  • Prioritise client operations and service reliability in all day-to-day activities and decision-making.

  • Undertake management responsibilities in line with company policies and applicable legislation, including interviewing, hiring and training employees.

  • Plan, allocate and oversee work for team members, support performance management activities and contribute to employee development.

  • Assist with resolving team challenges and fostering a positive, collaborative working environment.

  • Carry out additional duties as required to support the reliability, availability and performance of production systems.


Essential Skills


  • 5+ years' experience troubleshooting applications within production or operational environments.

  • 3+ years' experience writing and supporting SQL, with strong SQL development skills.

  • Good understanding of Microsoft SQL Server and SQL Server Management Studio (SSMS).

  • 3-5+ years' experience in Systems Engineering, including hands-on experience with Windows Server.

  • experience working within Agile environments using JIRA and SCRUM methodologies.

  • Proven ability to troubleshoot issues across application and hosting environments, including multi-tier architectures.

  • Enthusiasm for learning new technologies and the ability to apply new knowledge quickly.

  • Strong teamwork, communication and stakeholder management skills.

  • Good to advanced practical SQL programming experience.

  • Excellent problem-solving and analytical skills across both systems and applications.

  • Working knowledge of Windows operating systems within enterprise environments.


Additional Skills & Qualifications


  • Bachelor's degree in Engineering, Computer Science, or equivalent industry experience.

  • experience with job scheduling and workload management systems.

  • Basic object-oriented programming experience, ideally with C# and/or F#.

  • experience troubleshooting IT systems and infrastructure-related issues.

  • Exposure to AWS and cloud-based environments.

  • Knowledge of Windows and Windows Server administration.

  • Background supporting applications within 2nd and 3rd line support functions.

  • experience with software testing and QA best practices.

  • Familiarity with ECS, APIs and IIS Server environments.

  • Strong documentation skills, with the ability to clearly record processes, procedures and solutions.

  • Proactive mindset with the ability to identify and address issues before they impact services.

  • Ability to mentor colleagues while working collaboratively within team structures.


Location

London, UK


Rate/Salary

300.00 - 350.00 GBP Daily

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

DNS INFO LTD • City Of London

On-site
GBP 70,000 - 95,000
London-based SRE Ops & Reliability Engineer
London-based SRE Ops & Reliability Engineer

TEKsystems • Greater London

On-site
GBP 55,000 - 65,000
Application Support Analyst
Application Support Analyst

Euroclear • Greater London

On-site
GBP 45,000 - 62,000
Site Reliability Engineer - NS London
Site Reliability Engineer - NS London

BAE Systems Digital Intelligence • Greater London

Hybrid
GBP 50,000 - 70,000
Hybrid working environment
On-call allowances
Overtime benefits for night shifts
Site Reliability Engineer – NS London
Site Reliability Engineer – NS London

BAE Systems • Greater London

Hybrid
GBP 45,000 - 70,000
Hybrid working flexibility
On-call allowances
Overtime benefits
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Xpertise Recruitment • West Drayton

On-site
GBP 60,000 - 80,000
SRE Technical Lead
SRE Technical Lead

83zero Ltd • Wokingham

Hybrid
GBP 60,000 - 100,000
5% bonus
Hybrid working model
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Tenth Revolution Group • Knutsford

Hybrid
GBP 70,000 - 90,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Operations Engineer
Operations Engineer

AGS • United Kingdom

Hybrid
GBP 51,000 - 69,000