SRE Engineer

Vastbouw

Greater London

Hybrid

GBP 90,000 - 130,000

Full time

27 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Ricoh is seeking a Site Reliability Engineer to drive reliability, performance and operational excellence across hybrid cloud and on‑prem environments. You will shape SRE practices, support incidents, embed automation and ensure ISO 27001 security alignment.

This hands‑on role has significant influence across engineering, security and operations teams, with a focus on automation, observability and end‑to‑end service reliability.

Qualifications

  • Proven SRE/Production Engineering experience in medium to large environments.
  • Strong Azure and on‑prem infrastructure knowledge (IaaS, PaaS, networking, identity, storage).
  • Hands‑on IaC with Terraform or ARM/Bicep; CM with Ansible; CI/CD with Azure DevOps or GitHub Actions.
  • Experience with monitoring and observability stacks and security fundamentals.
  • Scripting or development ability (PowerShell, Python, Go).
  • Experience with containers and orchestration (Docker, Kubernetes, AKS).
  • Familiarity with ITSM platforms like ServiceNow and ISO 27001 environments.

Responsibilities

  • Deliver reliability against SLIs/SLOs; manage error budgets for core services.
  • Define standards for availability, latency, capacity, and scalability.
  • Lead root‑cause analyses and post‑mortem culture with actionable follow‑ups.
  • Embed observability with metrics, logs, traces and alerting.
  • Drive infrastructure‑as‑code and automation across Azure and on‑prem.
  • Support incident/problem management and security/compliance standards.

Skills

Azure
On‑prem infrastructure
Terraform
ARM/Bicep
Ansible
PowerShell DSC
CI/CD tooling
Monitoring
Windows OS
Linux OS
Scripting
PowerShell
Python
Go
Docker
Kubernetes
AKS
ITSM (ServiceNow)
ISO 27001 experience
Security
Communication

Tools

Terraform
ARM/Bicep
Ansible
PowerShell DSC
Azure DevOps
GitHub Actions
Monitoring stacks
Docker
Kubernetes
AKS
ServiceNow
ISO 27001 tooling

Job description

About Ricoh

A global leader in digital services, recognised for innovation, sustainability and a people-first culture. We feature in the Gartner Magic Quadrant, are listed in the Global 100 Most Sustainable Companies, and have been named one of Forbes’ World’s Best Employers 2025.

At Ricoh, we believe people do their best work when they feel valued and supported. We create inclusive workplaces where you can grow, contribute, and make a positive impact while helping to build a more sustainable future.

Find your place. Transform your future

Our purpose is centred on understanding and improving how people work. By focusing on real working experiences, we support individuals to develop their skills, realise their potential and do work that feels meaningful.

People transform when they Love What They Do

This belief sits at the heart of The Ricoh Promise. It guides how we recruit, how we support our people, and how we work together every day, creating an environment where you can grow, feel valued and make a difference.

When you join us, you are encouraged to share your ideas, challenge the way things are done, and work with others to build something better. If you are looking for a place where your voice is heard, your development is supported, and your work feels meaningful, you will feel at home at Ricoh.

What you will be doing

As a Site Reliability Engineer, you will play a key role in driving reliability, performance, and operational excellence across Ricoh’s hybrid cloud and on‑prem environments. You will help shape SRE practices, support incident and problem management, embed automation, and ensure infrastructure operations meet the highest security and compliance standards, including ISO 27001.

This is a hands‑on technical role with significant influence across engineering, architecture, security, and operational teams.

Responsibilities Include:

  • Delivering against SLIs, SLOs and managing error budgets for core services
  • Implementing standards for availability, latency, performance, capacity, and scalability
  • Leading and contributing to root‑cause analysis and major incident reviews
  • Supporting a blameless post‑mortem culture with clear action tracking
  • Defining and implementing SRE practices, tooling, and engineering standards
  • Driving infrastructure‑as‑code and automation across Azure and on‑prem
  • Improving image bakery pipelines for secure, repeatable server builds
  • Embedding observability using metrics, logs, traces, and effective alerting
  • Ensuring all practices align with ISO 27001 and internal security frameworks
  • Managing automated patching, vulnerability remediation and configuration compliance
  • Building dashboards and KPIs for reliability, MTTR, change failure rate, capacity and operational trends
  • Reducing operational toil through automation and improved tooling
  • Supporting the delivery and evolution of the SRE roadmap aligned to Ricoh’s transformation strategy
You will ideally have

Technical Skills & Experience

  • Proven experience in Site Reliability Engineering or Production Engineering within medium to large‑scale environments
  • Strong knowledge of Azure and on‑prem infrastructure (IaaS, PaaS, networking, identity, storage)
  • Hands‑on experience with infrastructure‑as‑code (Terraform, ARM/Bicep), configuration management (Ansible, PowerShell DSC), and CI/CD tooling (Azure DevOps, GitHub Actions)
  • Experience with monitoring and observability stacks
  • Solid understanding of OS fundamentals (Windows/Linux), security, networking
  • Background in scripting or software development (PowerShell, Python, Go)
  • Experience with containers and orchestration (Docker, Kubernetes, AKS)
  • Familiarity with ITSM practices and platforms such as ServiceNow
  • Experience operating in ISO 27001 or similar regulated environments

Business & Interpersonal Skills

  • Experience contributing to incident response, root‑cause analysis and continuous improvement
  • Ability to influence without direct authority and challenge the status quo
  • Comfortable working with security, architecture, product and operational teams
  • Strong communication skills with the ability to translate complex technical issues to non‑technical audiences
  • Calm and effective in high‑pressure situations, especially major incidents
In return for your commitment, you can expect

At Ricoh, work should feel meaningful, supportive and fulfilling. The Ricoh Promise shapes your experience through four pillars that bring our culture to life.

Love to Connect

You become part of a global community built on openness, inclusion and genuine collaboration.
Across teams, countries and roles, you'll find people who listen, involve and encourage you - helping you feel valued and able to be yourself every day.

Love to Grow

Your development truly matters to us. With access to learning pathways, mentoring and career opportunities across functions and countries, you'll be supported to stretch your skills, explore new directions and stay future-ready in a changing world.

Love to Give Back

Purpose is part of how we work. You'll have opportunities to make a difference through volunteering, sustainability initiatives and community programmes that reflect our shared values and commitment to positive impact.

Love to Succeed

Success at Ricoh is something we pursue together. You'll benefit from fair rewards, flexible working, wellbeing resources and real recognition - including programmes such as the Imagine. Change. Awards, where colleagues celebrate each other's achievements.

We are an equal opportunities employer

We believe that diverse perspectives make us stronger, and we welcome applications from people of all backgrounds, identities, and experiences. Our hiring decisions are based on skills, experience and potential, and we are committed to creating a fair and inclusive recruitment process. If you require any reasonable adjustments at any stage of the recruitment journey, please let us know and we will support you to bring your best self forward.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Platform Reliability Engineer
Senior Platform Reliability Engineer

Vastbouw • Greater London

On-site
GBP 90,000 - 130,000
Customer Support Technician
Customer Support Technician

Vastbouw • Reading

Hybrid
GBP 30,000 - 42,000
Site Services Supervisor
Site Services Supervisor

Vastbouw • West of England

On-site
GBP 24,000 - 30,000
Love to Connect
Love to Grow
Love to Give Back
+1
Site Services Supervisor
Site Services Supervisor

Ricoh UK • West of England

On-site
GBP 26,000 - 32,000
Flexible working
Site Services Manager - Print Room
Site Services Manager - Print Room

Ricoh UK • Nottingham

On-site
GBP 25,000 - 35,000
Career development opportunities
Volunteering initiatives
Flexible working hours
Service Request Analyst (Fixed Term) Entry Level
Service Request Analyst (Fixed Term) Entry Level

Vastbouw • Leeds

Hybrid
GBP 26,000 - 34,000
Service Request Analyst (Fixed Term) Entry Level
Service Request Analyst (Fixed Term) Entry Level

Ricoh • Leeds

On-site
GBP 23,000 - 35,000
Health & Safety Advisor
Health & Safety Advisor

Ricoh • Northampton

On-site
GBP 42,000 - 64,000
Flexible working
Wellbeing resources
Career progression
+1
IT Security Governance, Risk & Controls Analyst
IT Security Governance, Risk & Controls Analyst

Ricoh • City Of London

On-site
GBP 65,000 - 90,000
Senior Platform Reliability Engineer
Senior Platform Reliability Engineer

Ricoh • Greater London

On-site
GBP 62,000 - 102,000