Senior Site Reliability Engineer: Kubernetes & Automation Lead

EMBL-EBI

Hinxton

On-site

GBP 68,000 - 83,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Generous time off
Private medical insurance
Relocation package

Job summary

EMBL-EBI in Hinxton, UK, is seeking a skilled Site Reliability Engineer to design, deploy, and operate a Kubernetes-based web hosting platform for public-facing services. You will implement monitoring with Prometheus, oversee ElasticSearch analytics, and drive production automation using GitLab CI/CD.

The role requires strong Linux administration, Python scripting, and collaboration within a team to ensure reliable, scalable services. Relocation support and generous campus benefits are provided.

Qualifications

  • Bachelor's degree or higher in computer science or related discipline or equivalent experience.
  • At least 3 years of experience designing, implementing and operating large-scale web hosting platforms.
  • Experience managing public-facing production services.
  • At least 3 years of experience with automated deployment/configuration methods (e.g., Ansible, Puppet, Terraform).
  • Solid experience in Kubernetes deployment and administration in public or private cloud.
  • Strong Linux administration skills, ideally with RHEL or a RHEL clone.
  • Solid skills in automation tools like Jenkins, Rundeck, or similar.
  • Hands-on experience using Git in CI/CD and infrastructure-as-code workflows.
  • Solid skills in at least one programming language, ideally Python.
  • Experience with infrastructure monitoring methodologies.
  • Strong interpersonal and written English communication skills.
  • Proven ability to work well in a team, building positive relationships and sharing knowledge.
  • Ability to plan and prioritise workloads.

Responsibilities

  • Build and maintain the web hosting platform based on Kubernetes, allowing users to deploy web applications.
  • Implement and manage infrastructure and application monitoring based on Prometheus.
  • Oversee the web analytics platform currently based on ElasticSearch.
  • Utilize CI/CD tools like Gitlab for automation and operational efficiency.
  • Ensure documentation meets standard requirements.
  • Drive SRE best practices throughout the team.
  • Contribute directly to projects and tasks related to production automation.
  • Assist and guide team members with the daily prioritisation of tasks.

Skills

Automation tooling
CI/CD
Git in CI/CD
Python
Kubernetes
Linux administration
Monitoring
Teamwork
Work prioritisation

Education

Bachelor's degree or higher in CS or related

Tools

Ansible
Puppet
Terraform
Kubernetes
Jenkins
Rundeck
Git
GitLab
Prometheus
ElasticSearch

Job description

EMBL-EBI in Hinxton, UK, is seeking a skilled Site Reliability Engineer to design, deploy, and operate a Kubernetes-based web hosting platform for public-facing services. You will implement monitoring with Prometheus, oversee ElasticSearch analytics, and drive production automation using GitLab CI/CD.

The role requires strong Linux administration, Python scripting, and collaboration within a team to ensure reliable, scalable services. Relocation support and generous campus benefits are provided.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

EMBL-EBI • Hinxton

On-site
GBP 68,000 - 83,000
Generous time off
Private medical insurance
Relocation package
Senior Site Reliability Engineer - Research IT Ops (Hybrid)
Senior Site Reliability Engineer - Research IT Ops (Hybrid)

European Molecular Biology Laboratory • Hinxton

Hybrid
GBP 43,000 - 54,000
Hybrid working
Relocation package
Private medical insurance
+3
Senior SRE: Global Research Platform (Hybrid)
Senior SRE: Global Research Platform (Hybrid)

European Bioinformatics Institute • Cambridge

Hybrid
GBP 55,000 - 75,000
Private medical insurance
30 days annual leave
Hybrid working pattern
+2
Site Reliability Engineer - AWS, Kubernetes, CI/CD (EMEA)
Site Reliability Engineer - AWS, Kubernetes, CI/CD (EMEA)

zerohash • Greater London

On-site
GBP 90,000 - 150,000
Equity
Maternity & Paternity leave
WeWork Membership
+2
Senior SRE — Hybrid Global Research Infrastructure
Senior SRE — Hybrid Global Research Infrastructure

EMBL • Hinxton

Hybrid
GBP 42,000 - 48,000
Hybrid work
Relocation package
Private medical insurance
+3
Kubernetes SRE: DevOps, Observability & Reliability
Kubernetes SRE: DevOps, Observability & Reliability

The Consensus • Greater London

Hybrid
GBP 85,000 - 125,000
SRE Kubernetes | Hybrid/WFH, Growth & Reliability Impact
SRE Kubernetes | Hybrid/WFH, Growth & Reliability Impact

Client Server • Cambridge

Hybrid
GBP 59,000 - 81,000
Pension
Private Medical Insurance
Life Assurance
+5
Senior SRE: GCP & Kubernetes, Automation Lead
Senior SRE: GCP & Kubernetes, Automation Lead

P2P • Greater London

On-site
GBP 90,000 - 130,000
EMEA Site Reliability Engineer - Cloud, CI/CD & Infra
EMEA Site Reliability Engineer - Cloud, CI/CD & Infra

Zero Hash • United Kingdom

Hybrid
GBP 90,000 - 140,000
Equity opportunity
Maternity & Paternity leave
WeWork Membership
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Selby Jennings • Greater London

On-site
GBP 70,000 - 90,000