Senior Site Reliability Engineer

Spectrum IT Recruitment

Southampton

Hybrid

GBP 55,000 - 90,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Life Insurance
Private Medical Insurance
Employee Assistance Programme
Hybrid working
GP Online Portal
Plus More

Job summary

Spectrum IT Recruitment is seeking an experienced DevOps/SRE Engineer for a Southampton-based role, with a hybrid model (3 days in office). You will oversee production environments, ensuring availability, building tooling for platform infrastructure, and accelerating delivery speed across distributed applications.

The role emphasizes performance tuning, proactive monitoring, and collaboration with developers to improve service quality, resilience, and release practices.

Qualifications

  • Experience in building and operating scalable cloud-based platforms.
  • Proficiency with observability tooling and incident management.
  • Ability to design automated, reliable infrastructure and CI/CD pipelines.

Responsibilities

  • Monitor and fine-tune system performance to meet demand.
  • Collaborate with developers to improve service quality via testing and structured releases.
  • Lead capacity forecasting and manage platform operations across distributed apps.
  • Design automated solutions to build resilient, scalable systems.
  • Provide operational support and technical oversight for large-scale applications.

Skills

CI/CD fundamentals
Cloud platforms
Systems engineering
Automation scripting

Tools

Kubernetes
Grafana
Splunk
Datadog
PagerDuty
Rundeck
Ansible
Puppet
Chef
AWS DevOps

Job description

Southampton HQ - 2 Times a week in Office
Cloud, SaaS, AWS,

The company deliver cutting-edge enterprise software solutions across both cloud and on-premises environments, empowering organisations to enhance customer experiences, maintain regulatory compliance, and proactively fight fraud. The company are trusted by businesses worldwide to drive seamless, intelligent customer interactions.

In this role, you’ll oversee the production environment by ensuring system availability and maintaining a comprehensive perspective on overall health. You’ll develop tools and software to support and streamline the management of platform infrastructure and key applications. A major focus will be enhancing the dependability, performance, and delivery speed of our software products. You’ll also be responsible for analysing and fine-tuning system performance to anticipate user demands and drive innovation. Additionally, you’ll take the lead in providing operational support and technical oversight for several large-scale distributed applications.

You’ll Contribute:
  • Monitor and interpret system and application metrics to fine-tune performance and troubleshoot issues effectively
  • Collaborate closely with developers to enhance service quality through thorough testing and structured release practices
  • Engage in architectural discussions, manage platform operations, and contribute to capacity forecasting
  • Design and implement automated solutions to build resilient, scalable systems
  • Maintain a strong focus on delivering new features while ensuring stability and adherence to service level goals
You’ll Stand Out If You Have:
  • Practical experience managing large-scale Kubernetes clusters; certifications in Kubernetes are a strong bonus
  • Hands-on familiarity with the Grafana Observability Suite, including tools like Loki, Mimir, and Tempo
  • Background in administering or developing with popular monitoring and automation tools such as Splunk, Datadog, PagerDuty, or Rundeck
  • Experience using configuration management platforms like Ansible, Puppet, or Chef
  • Professional certifications in cloud DevOps, such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer, or similar credentials
Do You Have What It Takes?
  • 3-6 years of hands-on experience in a similar role, with a strong emphasis on systems engineering, automation, and service reliability
  • Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills in Bash or PowerShell
  • Solid grasp of cloud platforms like AWS, including an understanding of how core services like EC2, ECS, Lambda, and DynamoDB operate under reliability constraints
  • Practical experience using infrastructure-as-code tools like CloudFormation or Terraform
  • In-depth knowledge of CI/CD principles and hands-on experience with tools such as Jenkins, GitLab CI/CD, or CircleCI
  • Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture
  • Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch
  • Excellent analytical and troubleshooting abilities, especially within complex distributed systems
  • Proven experience handling incident management and conducting blameless postmortems, including leading cross-functional teams through resolution and communication during critical outages
  • Life Insurance - 4 x Annual Salary
  • Private Medical Insurance
  • Employee Assistance Programme
  • Hybrid Working - 3 Days from Home
  • GP Online Assistance Portal.
  • + Much More
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AWS Site Reliability Engineer
Senior AWS Site Reliability Engineer

Spectrum IT Recruitment • City Of London

Hybrid
GBP 65,000 - 120,000
Life Insurance 4x Annual Salary
Private Medical Insurance
Bonus Scheme
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Symphony • Belfast City District

On-site
GBP 60,000 - 70,000
Regional specific competitive benefits
Build your own Benefits (BYOB) perk
Local events, team building, and devop
Lead Site Reliability Engineer - Glasgow
Lead Site Reliability Engineer - Glasgow

Hackajob Ltd • Glasgow

On-site
GBP 90,000 - 110,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Square One Resources • Sutton Coldfield

Hybrid
Service Manager – Site Reliability Engineering
Service Manager – Site Reliability Engineering

Jobtailor • Belfast City District

On-site
GBP 85,000 - 110,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Omilia • Greater London

On-site
GBP 90,000 - 120,000
Fixed compensation
Long-term vacation
Professional growth
+3
Senior DevOps Engineer
Senior DevOps Engineer

Sopra Steria Ltd • Cheltenham

On-site
GBP 65,000 - 90,000
Car allowance
Annual leave
Private medical
+3
Director of Site Reliability Engineering
Director of Site Reliability Engineering

EPAM Systems • Greater London

Hybrid
GBP 180,000 - 240,000
ESPP
Life Assurance
Income protection
+14
Site Reliability Engineer
Site Reliability Engineer

ReVybe IT Recruitment Limited • Greater London

Hybrid
GBP 51,000 - 85,000
Bonus
Benefits