Senior Platform Reliability Engineer

Ricoh

Greater London

On-site

GBP 62,000 - 102,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Ricoh in London seeks a Senior Platform Reliability Engineer to own the reliability, resilience, and operational integrity of our hybrid and cloud platform environment.

You will drive infrastructure as code and automation (Azure, Terraform, Ansible), embed observability, and support incident and problem management with SRE and helpdesk teams. Expect KPIs, dashboards, and cost-awareness to shape decisions. Occasional travel to datacentres may be required; sponsorship is not available.

Qualifications

  • Senior hands-on engineering with cloud, infra and operations.
  • Experience balancing cloud, cost, risk, security and reliability.
  • Strong Azure iaaS/paas and on‑prem data centre operations.

Responsibilities

  • Deliver availability, latency, performance, capacity and scalability standards.
  • Lead root-cause analysis for major incidents and drive post-mortems.
  • Drive infrastructure as code and automation across Azure and co-location environments.
  • Evolve image bakery pipelines for secure, repeatable server images.
  • Embed observability with metrics, logs, traces and alerting tools.
  • Partner with SRE and helpdesk teams to deliver reliable service.
  • Oversee automated patching, vulnerability remediation and config compliance.
  • Introduce KPIs and dashboards for reliability, incidents and capacity.
  • Travel to datacentres and offices to ensure projects meet business requirements.

Skills

Business awareness
Incident management
Communication
Service excellence

Tools

Azure
Terraform
Ansible
PowerShell
GitHub Actions
CI/CD
ServiceNow
ITIL
ARM/Bicep
Linux
Windows
Observability

Job description

Salary: £62,000 - 102,000 per year

Requirements:
  • We are looking for a senior, hands-on engineering professional with strong practical knowledge of infrastructure, cloud, and operations beyond task execution.
  • We require strong business awareness and the ability to understand how infrastructure services support business operations, customer experience, and strategic objectives within a Cloud-First strategy.
  • We need experience making decisions that balance cloud and managed services, technical quality, cost efficiency, risk, security, and service reliability.
  • We are looking for previous experience in a similar role.
  • We require strong background in Azure, including IaaS, PaaS, networking, identity, and storage, as well as on-prem data centre operations.
  • We need hands-on skills with infrastructure as code and automation, including Terraform, ARM/Bicep, Ansible, PowerShell DSC, and CI/CD pipelines such as Azure DevOps and GitHub Actions.
  • We require experience with monitoring and observability tools, alert design, and dashboarding.
  • We value commercial awareness, including experience working with vendors and partners, understanding cloud consumption models, licensing and support contracts, and contributing to cost optimisation, business cases, and risk assessments across cloud and hybrid environments.
  • We require a service-focused mindset, with knowledge of IT service management, clear communication and documentation, effective incident and problem support, and continuous improvement of platforms and services.
  • We require knowledge of networking, security, and operating system fundamentals for Windows and Linux.
  • We prefer experience operating in ISO 27001 or similar regulated environments.
  • We prefer experience integrating with ITSM platforms such as ServiceNow and aligning with ITIL processes.
  • We require an understanding of the business and application impact of infrastructure decisions.
  • We need working knowledge of security, architecture, and vendor/commercial considerations to support informed decision-making.
Responsibilities:
  • Deliver standards for availability, latency, performance, capacity, and scalability.
  • Take part in root-cause analysis and problem management for major incidents.
  • Champion a blameless post-mortem culture and ensure actions are tracked and closed.
  • Drive infrastructure as code and automation across Azure and co-lo environments.
  • Evolve the image bakery pipeline for secure, repeatable server images.
  • Embed observability using metrics, logs, traces, and alerting tools.
  • Partner with SRE and helpdesk teams to deliver service.
  • Oversee automated patching, vulnerability remediation, and configuration compliance.
  • Introduce KPIs and dashboards for reliability, incident trends, MTTR, change failure rate, and capacity.
  • Travel occasionally to datacentres and offices to ensure projects and services meet business requirements.
Technologies:
  • ARM
  • Ansible
  • Azure
  • CI/CD
  • Cloud
  • DevOps
  • GitHub
  • IaaS
  • Support
  • ITIL
  • ITSM
  • Linux
  • PaaS
  • PowerShell
  • Security
  • ServiceNow
  • Terraform
  • Windows
More:

We are Ricoh, a global leader in digital services recognised for innovation, sustainability, and a people-first culture. We are listed in the Gartner Magic Quadrant, the Global 100 Most Sustainable Companies, and Forbes Worlds Best Employers 2025. We believe people do their best work when they feel valued and supported, and we create inclusive workplaces where you can grow, contribute, and make a positive impact while helping to build a more sustainable future. We are recruiting for a Senior Platform Reliability Engineer based in London. This is a senior, hands-on role focused on the reliability, resilience, and operational integrity of our hybrid and cloud platforms. We are unable to provide sponsorship for this position. In return, we offer the Ricoh Promise, which includes opportunities to connect globally, grow through learning and mentoring, give back through volunteering and sustainability initiatives, and succeed with fair rewards, flexible working, wellbeing resources, and recognition. We are an equal opportunities employer and welcome applicants from all backgrounds, identities, and experiences.

last updated 36 week of 2026

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Reliability Engineer – Azure, IaC & Observability
Senior Cloud Reliability Engineer – Azure, IaC & Observability

Ricoh • Greater London

On-site
GBP 62,000 - 102,000
Senior Engineer
Senior Engineer

Back TO Work • City of Westminster

On-site
GBP 90,000 - 100,000
Cloud Engineer
Cloud Engineer

Arthur Recruitment • Greater London

Hybrid
GBP 100,000 - 150,000
Cloud Operations Service Reliability Engineer
Cloud Operations Service Reliability Engineer

A&O Shearman • Carrickfergus

On-site
GBP 45,000 - 73,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

NICE Systems • Southampton

Hybrid
GBP 40,000 - 60,000
Principal SRE (AWS, Azure, Terraform, Kubernetes)
Principal SRE (AWS, Azure, Terraform, Kubernetes)

Fourth • Greater London

Hybrid
GBP 50,000 - 90,000
Hybrid working
Pension and life insurance
Healthcare expense claims
+4
Senior Systems Engineer - Secure Cloud & Infrastructure
Senior Systems Engineer - Secure Cloud & Infrastructure

CBSbutler Holdings Limited trading as CBSbutler • Reading

Hybrid
GBP 70,000 - 100,000
Lead Site Reliability Engineer - Edinburgh
Lead Site Reliability Engineer - Edinburgh

Inspire People • City of Edinburgh

Hybrid
GBP 72,000 - 88,000
Platform Engineer
Platform Engineer

System C • Brentwood

On-site
GBP 40,000 - 50,000
Flexible benefits
Private healthcare
Pension options
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VIQU IT Recruitment • Milton Keynes

On-site
GBP 45,000 - 75,000