Site Reliability Engineer

Royal Society of Chemistry

Cambridge

On-site

GBP 58,000 - 71,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

The Royal Society of Chemistry in Cambridge is seeking a Site Reliability Engineer to join our DevOps team. You will design, deploy and operate AWS-based cloud services to improve reliability, availability and performance across staff tools, websites and data platforms.

You will advance observability, CI/CD, automation and secure data delivery using S3, DataSync, CloudFront and Transfer Family, collaborating with product teams and data owners to deliver scalable, resilient solutions.

Qualifications

  • Experience designing, deploying and operating AWS infrastructure.
  • Familiar with Terraform, CI/CD, and multi-environment deployments.
  • Strong collaboration with product teams and data owners.

Responsibilities

  • Deliver reliable and efficient services using SRE principles, with observability and SLIs/SLOs.
  • Automate toil, improve self-service and streamline deployments.
  • Design and maintain monitoring and alerting for service health.
  • Review tools and pipelines to ensure secure, scalable DevOps delivery.
  • Lead post-incident reviews and implement corrective actions.

Skills

AWS cloud
CI/CD
Observability
Team collaboration

Education

Tools

Terraform
AWS
DataSync
S3
CloudFront
Transfer Family

Job description

Circa: Salary - Salary Plan, 64,680.00 GBP Annual Are you passionate about building reliable, scalable cloud services and solving complex operational challenges? We're looking for a Site Reliability Engineer to join our DevOps team and play a vital role in shaping the future of our technology platforms. You'll help ensure the reliability, availability and performance of the services that support our staff, websites, digital products, data services and customer-facing platforms, enabling the Royal Society of Chemistry to deliver an excellent experience for users across the globe. This is an exciting opportunity to contribute to a significant data-focused project, where you will help design and deliver modern cloud solutions that support the transfer, processing, storage and distribution of large-scale datasets. Working with product teams, data owners and technical colleagues, you will help develop robust AWS-based architectures that enable secure and efficient movement of data, leveraging technologies such as Amazon S3, automated processing pipelines and scalable distribution services. Your work will directly support the delivery of critical information and digital services to a wide range of audiences. We're looking for a technically strong and customer-focused problem solver with experience designing and operating AWS cloud solutions at scale. You'll have a solid understanding of DevOps practices, including continuous integration, continuous deployment, configuration management and process automation, alongside a passion for improving how services are delivered and supported. Most importantly, you'll enjoy working collaboratively with stakeholders to create reliable, secure and efficient solutions that make a real impact. If you're excited by the challenge of combining cloud engineering, Site Reliability Engineering and large-scale data delivery in a purpose-driven organisation, we'd love to hear from you.

Day to day activities:

Deliver reliable and efficient services by applying Site Reliability Engineering principles, including observability, automation, service level objectives and continuous improvement. Reduce operational burden by automating manual work, reducing toil, improving self-service capabilities and supporting continuous process improvement. Design, implement and maintain monitoring and observability capabilities that support service health measurement, alerting, SLIs and SLOs. Support secure and effective DevOps delivery by reviewing and validating tools, pipelines, infrastructure, platform services and environment provisioning. Manage operational change effectively through service transitions, roll-outs, configuration changes and decommissioning activity. Improve service resilience after incidents by investigating and resolving incidents, problems, risks and reliability issues, and by contributing to blameless post-incident reviews that lead to prioritised corrective action.

Requirements:

A Demonstrable experience designing, deploying and operating AWS infrastructure using Terraform, including reusable modules, remote state management, CI/CD integration, code reviews and multi-environment deployments. Experience designing and supporting large-scale data delivery solutions using AWS services such as S3, DataSync, Transfer Family, CloudFront, Direct Connect and cross-account access patterns. Understanding of data transfer performance, throughput optimisation, storage lifecycle management, integrity validation and customer-facing data delivery. Knowledge of DevOps concepts, including continuous integration, continuous deployment, configuration management and process automation. Knowledge of software security vulnerabilities and mitigation. Ability to collaborate with product teams, data owners and customers, providing clear technical advice and supporting successful delivery requirements.

Visit our Work For Us website to learn more about us, our benefits, equal opportunities statement and inclusive culture pledge.

At the RSC, we recognise the benefits of a diverse workforce and welcome applicants from a range of backgrounds to apply. We particularly encourage applications from disabled and ethnic minority candidates. As a part of the Disability Confident Scheme, we endeavour, where possible, to offer an interview to candidates meeting the essential criteria of the role, who has a substantial physical/mental impairment which impacts their ability to carry out day-to-day tasks. We are committed to making our recruitment processes accessible to all and as part of this, we are flexible in the ways we give and receive information.

Being part of the Royal Society of Chemistry is to be at the heart of a dedicated, forward-thinking and passionate collection of people. You’ll find opportunities to work in publishing, membership, sales, marketing, communications, technology and finance, community outreach, and many others. Everything we do, no matter the area, is to support a global scientific community. The groups we work with are exceptionally diverse – from teachers, scientists and academics, librarians and corporates to politicians and the public.

Thank you for your interest in joining the Royal Society of Chemistry!

AI tools can be a great support when putting together your application materials.

Guidance on the appropriate use of AI tools

We are a Disability Confident Committed Employer

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer — AWS Cloud & Data Delivery
Site Reliability Engineer — AWS Cloud & Data Delivery

Royal Society of Chemistry • Cambridge

On-site
GBP 58,000 - 71,000
Software Development Team Lead
Software Development Team Lead

Royal Society of Chemistry • Cambridge

Hybrid
GBP 53,000 - 73,000
Lead Site Reliability Engineer (AWS)
Lead Site Reliability Engineer (AWS)

United States Digital Space LLC • Greater London

Hybrid
GBP 95,000 - 130,000
Hybrid working
Site Reliability Engineer
Site Reliability Engineer

Wedo Technology Solutions Ltd. • Greater London

Remote
GBP 63,000 - 75,000
HRIS Consultant
HRIS Consultant

Royal Society of Chemistry • Cambridge

On-site
GBP 70,000 - 90,000
Site Reliability Engineer
Site Reliability Engineer

SYNALOGiK Innovative Solutions Limited • Hereford

Hybrid
GBP 45,000 - 55,000
Private medical insurance
Dental insurance
Pension scheme
+2
Site Reliability Engineer - Negotiable
Site Reliability Engineer - Negotiable

Alchemy • Reading

Hybrid
GBP 60,000 - 80,000
Competitive salary
Healthcare benefits
Site Reliability Engineer
Site Reliability Engineer

Lloyds Bank plc • City of Westminster

Hybrid
GBP 84,000 - 93,000
Director of Site Reliability Engineering
Director of Site Reliability Engineering

EPAM Systems • Greater London

Hybrid
GBP 140,000 - 200,000
ESPP
Life assurance
Income protection
+11
Site Reliability Engineer – NS London
Site Reliability Engineer – NS London

BAE Systems • Greater London

Hybrid
GBP 45,000 - 70,000
Hybrid working flexibility
On-call allowances
Overtime benefits