Senior Site Reliability Engineer

Spectrum IT Recruitment

Southampton

Hybrid

GBP 70,000 - 110,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Life Insurance 4x salary
Private Medical Insurance
Employee Assistance Programme
Hybrid Working 3 days from Home
GP Online Assistance Portal

Job summary

Spectrum IT Recruitment는 Southampton에 위치한 본사에서 Production 환경을 감독하고 가용성을 유지하며 플랫폼 인프라 관리 도구를 개발하는 DevOps/클라우드 엔지니어를 찾습니다. 2회 주간 사무실 근무를 포함하는 하이브리드 근무 형태로, Kubernetes, AWS, CI/CD, 모니터링 도구에 능숙한 지원자를 환영합니다.

관계 부서와 협력하여 서비스 품질을 개선하고 아키텍처 논의에 참여하며 자동화 솔루션을 설계합니다. 대규모 분산 시스템의 장애 대응 및 포스트모트 진행 경험이 큰 강점이 됩니다.

Qualifications

  • 3–6년의 실무 경험, 대규모 분산 시스템에서의 운영 및 자동화, 서비스 가용성에 중점
  • Python/Go/Java/C# 중 하나 이상의 프로그래밍 언어 숙련 및 Bash/PowerShell 스크립트 능력
  • AWS를 포함한 클라우드 플랫폼 및 EC2/ECS/Lambda/DynamoDB의 안정성 제어 이해
  • Terraform/CloudFormation 등 인프라 자동화 도구 사용 경험
  • CI/CD 원칙에 대한 확고한 이해와 Jenkins/GitLab CI/CD/CircleCI 등의 실무 경험
  • Docker 및 Kubernetes 기반 컨테이너화와 마이크로서비스 아키텍처에 대한 실무 지식
  • Prometheus/Grafana/ELK/AWS CloudWatch 등의 관찰성 도구 사용 경험
  • 사건 관리 및 교차 기능 팀 리더십 경험(블레이멈 포스트모트 포함)

Responsibilities

  • 시스템 및 애플리케이션 지표 모니터링으로 성능 미세조정 및 문제 해결
  • 개발자와 협업하여 서비스 품질 향상 및 테스트/릴리스 프로세스 개선
  • 아키텍처 논의에 참여하고 플랫폼 운영 및 용량 예측 관리
  • 자동화 솔루션 설계 및 확장 가능한 시스템 구축
  • 서비스의 새로운 기능 배포와 가용성/서비스 수준 목표 달성에 집중

Skills

Kubernetes
Grafana
Splunk
Datadog
PagerDuty
Rundeck
Ansible
Puppet
Chef
AWS DevOps
Terraform
CloudFormation
Jenkins
GitLab CI
CircleCI
Docker
Prometheus
ELK
CloudWatch
Python
Go
Java
C#
Bash
PowerShell

Tools

Terraform
CloudFormation
Jenkins
GitLab CI/CD
CircleCI
Ansible
Puppet
Chef
Docker

Job description

Southampton HQ - 2 Times a week in Office
Cloud, SaaS, AWS,

The company deliver cutting-edge enterprise software solutions across both cloud and on-premises environments, empowering organisations to enhance customer experiences, maintain regulatory compliance, and proactively fight fraud. The company are trusted by businesses worldwide to drive seamless, intelligent customer interactions.

In this role, you'll oversee the production environment by ensuring system availability and maintaining a comprehensive perspective on overall health. You'll develop tools and software to support and streamline the management of platform infrastructure and key applications. A major focus will be enhancing the dependability, performance, and delivery speed of our software products. You'll also be responsible for analysing and fine-tuning system performance to anticipate user demands and drive innovation. Additionally, you'll take the lead in providing operational support and technical oversight for several large-scale distributed applications.

How You'll Contribute:
  • Monitor and interpret system and application metrics to fine-tune performance and troubleshoot issues effectively
  • Collaborate closely with developers to enhance service quality through thorough testing and structured release practices
  • Engage in architectural discussions, manage platform operations, and contribute to capacity forecasting
  • Design and implement automated solutions to build resilient, scalable systems
  • Maintain a strong focus on delivering new features while ensuring stability and adherence to service level goals
You'll Stand Out If You Have:
  • Practical experience managing large-scale Kubernetes clusters; certifications in Kubernetes are a strong bonus
  • Hands-on familiarity with the Grafana Observability Suite, including tools like Loki, Mimir, and Tempo
  • Background in administering or developing with popular monitoring and automation tools such as Splunk, Datadog, PagerDuty, or Rundeck
  • Experience using configuration management platforms like Ansible, Puppet, or Chef
  • Professional certifications in cloud DevOps, such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer, or similar credentials
Do You Have What It Takes?
  • 3-6 years of hands-on experience in a similar role, with a strong emphasis on systems engineering, automation, and service reliability
  • Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills in Bash or PowerShell
  • Solid grasp of cloud platforms like AWS, including an understanding of how core services like EC2, ECS, Lambda, and DynamoDB operate under reliability constraints
  • Practical experience using infrastructure-as-code tools like CloudFormation or Terraform
  • In-depth knowledge of CI/CD principles and hands-on experience with tools such as Jenkins, GitLab CI/CD, or CircleCI
  • Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture
  • Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch
  • Excellent analytical and troubleshooting abilities, especially within complex distributed systems
  • Proven experience handling incident management and conducting blameless postmortems, including leading cross-functional teams through resolution and communication during critical outages
  • Life Insurance - 4 x Annual Salary
  • Private Medical Insurance
  • Employee Assistance Programme
  • Hybrid Working - 3 Days from Home
  • GP Online Assistance Portal.
  • + Much More
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AWS Site Reliability Engineer
Senior AWS Site Reliability Engineer

Spectrum IT Recruitment • City Of London

Hybrid
GBP 85,000 - 110,000
Life Insurance - 4 x Annual Salary
Private Medical Insurance
Bonus Scheme
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Symphony • Belfast City District

On-site
GBP 60,000 - 70,000
Regional specific competitive benefits
Build your own Benefits (BYOB) perk
Local events, team building, and devop
Site Reliability Engineer
Site Reliability Engineer

ReVybe IT Recruitment Limited • Greater London

Hybrid
GBP 51,000 - 85,000
Bonus
Benefits
Site Reliability Engineer - Negotiable
Site Reliability Engineer - Negotiable

Alchemy • Reading

Hybrid
GBP 60,000 - 80,000
Competitive salary
Healthcare benefits
Senior DevOps Engineer
Senior DevOps Engineer

LaraBench • Manchester

Hybrid
GBP 55,000 - 65,000
Fully remote
Flexible hours
Bonus scheme
+2
Senior AWS Platform Engineer
Senior AWS Platform Engineer

ReVybe IT Recruitment Limited • City Of London

Hybrid
GBP 76,000 - 90,000
Hybrid work in London office (2 days)
Benefits
Senior Site Reliability Engineer
Senior Site Reliability Engineer

EMBL-EBI • Hinxton

On-site
GBP 68,000 - 83,000
Generous time off
Private medical insurance
Relocation package
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Square One Resources • Sutton Coldfield

Hybrid
Senior DevOps Systems Administrator
Senior DevOps Systems Administrator

Onyx-Conseil • Guildford

Hybrid
GBP 60,000 - 65,000
Private Medical Insurance
Holiday
Pension
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VIQU IT Recruitment • Kingston

On-site
GBP 68,000 - 83,000
On-call allowance
Bonus