London-based SRE Ops & Reliability Engineer

TEKsystems

Greater London

On-site

GBP 55,000 - 65,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TEKsystems in London, UK is seeking an experienced SRE Operations Support to act as a Subject Matter Expert within a Site Reliability Engineering team. You will manage, deploy, troubleshoot and monitor production systems to ensure reliability and performance across environments.

Working closely with support and engineering, you will implement automation, develop deployment guidelines, write SQL scripts, participate in incident management, and mentor teammates while maintaining documentation.

Qualifications

  • 5+ years' experience troubleshooting applications in production or operational environments.
  • 3+ years' experience writing and supporting SQL, with strong SQL development skills.
  • Proficient in Microsoft SQL Server and SSMS.
  • 3-5+ years' experience in Systems Engineering, including hands-on experience with Windows Server.
  • Experience working in Agile environments using JIRA and SCRUM methodologies.
  • Proven ability to troubleshoot issues across application and hosting environments, including multi-tier architectures.
  • Enthusiasm for learning new technologies and applying knowledge quickly.
  • Strong teamwork, communication and stakeholder management skills.
  • Good to advanced practical SQL programming experience.
  • Excellent problem-solving and analytical skills across systems and applications.
  • Working knowledge of Windows operating systems within enterprise environments.

Responsibilities

  • Deploy applications across all production environments, including Staging, Pre-Production, UAT and Production.
  • Troubleshoot application and system issues, identifying root causes and implementing effective solutions.
  • Perform testing and quality assurance activities following deployments to validate functionality and stability.
  • Develop and enhance testing and deployment guidelines to support consistent and reliable releases.
  • Monitor client systems to ensure services operate as expected and performance stays within thresholds.
  • Implement automation processes using proprietary worker and job management systems to improve efficiency.
  • Collaborate with support teams to triage, investigate and resolve issues impacting live systems.
  • Create SQL queries and scripts to process workloads and support operational requirements.
  • Investigate system performance issues and recommend optimisation opportunities.
  • Support day-to-day operation of systems through proactive monitoring and early issue detection.
  • Work with support teams on analysis, incident management and problem resolution.
  • Apply strong analytical and problem-solving skills when investigating issues.
  • Maintain a structured, disciplined approach to work and documentation.
  • Adhere to team processes and workflows, contributing to a collaborative environment.
  • Develop knowledge across systems and technologies, adapting quickly to new tools.
  • Communicate effectively with cross-functional teams when requesting or providing support.
  • Prioritise client operations and service reliability in daily activities and decisions.
  • Perform management responsibilities including interviewing, hiring and training employees.
  • Plan, allocate and oversee work for team members, support performance management and development.
  • Assist with resolving team challenges and fostering a positive, collaborative environment.
  • Carry out additional duties to support reliability, availability and performance of production systems.

Skills

SQL
SSMS
Windows Server
Agile
Troubleshooting
Automation
Monitoring
Incident Mgmt
Communication
Teamwork
C#
AWS

Education

Bachelor's degree in Engineering/Computer Science

Tools

JIRA
IIS
ECS
AWS

Job description

TEKsystems in London, UK is seeking an experienced SRE Operations Support to act as a Subject Matter Expert within a Site Reliability Engineering team. You will manage, deploy, troubleshoot and monitor production systems to ensure reliability and performance across environments.

Working closely with support and engineering, you will implement automation, develop deployment guidelines, write SQL scripts, participate in incident management, and mentor teammates while maintaining documentation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE Operations Support
SRE Operations Support

TEKsystems • Greater London

On-site
GBP 55,000 - 65,000
SRE Manager: Reliability & Incident Leadership (Hybrid London)
SRE Manager: Reliability & Incident Leadership (Hybrid London)

Gravitas Recruitment Group (Global) Ltd • Greater London

Hybrid
GBP 75,000 - 100,000
Site Reliability Engineer
Site Reliability Engineer

DNS INFO LTD • City Of London

On-site
GBP 70,000 - 95,000
Site Reliability Engineer — Production & Incident Response
Site Reliability Engineer — Production & Incident Response

慨正橡扯 • Greater London

Hybrid
GBP 60,000 - 80,000
SRE
SRE

Technopride Ltd • Hove

Hybrid
GBP 60,000 - 80,000
On-Site SRE: Cloud Reliability & Automation
On-Site SRE: Cloud Reliability & Automation

Talenzon group • Greater London

Hybrid
GBP 70,000 - 110,000
Site Reliability Engineer – NS London
Site Reliability Engineer – NS London

BAE Systems • Greater London

Hybrid
GBP 45,000 - 70,000
Hybrid working flexibility
On-call allowances
Overtime benefits
Site Reliability Engineer (SRE) – Cloud Platforms
Site Reliability Engineer (SRE) – Cloud Platforms

Talenzon group • Greater London

Hybrid
GBP 70,000 - 110,000
Site Reliability Engineer - NS London
Site Reliability Engineer - NS London

BAE Systems Digital Intelligence • Greater London

Hybrid
GBP 50,000 - 70,000
Hybrid working environment
On-call allowances
Overtime benefits for night shifts
Head of SRE & Reliability — Onsite in London
Head of SRE & Reliability — Onsite in London

Globant • Greater London

On-site
GBP 150,000 - 210,000