Site Reliability Engineer: On-Prem Kubernetes & MLOps

Helsing

City Of London

On-site

GBP 60,000 - 80,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

A defence AI company in the UK is seeking a Site Reliability Engineer to support high-security environments. In this role, you will design and manage Kubernetes infrastructure, ensuring system reliability through observability frameworks and collaboration with security teams. The ideal candidate has expertise in cloud-native technologies and scripting, contributing to impactful AI solutions. This is a full-time position at a mid-senior level.

Qualifications

  • Experience with cloud-native workloads in on-premises or air-gapped environments.
  • High level of personal integrity and attention to detail.
  • Software engineering mindset with a passion for productivity.

Responsibilities

  • Design, implement, and manage Kubernetes infrastructure.
  • Create observability frameworks using Grafana and Prometheus.
  • Collaborate with Security teams for supply chain security.

Skills

Scripting
GitOps workflows
Kubernetes expertise
Cloud-native technologies
Observability stack
Networking concepts
MLOps platforms
Infrastructure as code
System administration
Data and telemetry pipelines

Tools

Terraform
Ansible
Grafana
Prometheus
Kubeflow
Helm
Istio
OpenTelemetry

Job description

A defence AI company in the UK is seeking a Site Reliability Engineer to support high-security environments. In this role, you will design and manage Kubernetes infrastructure, ensuring system reliability through observability frameworks and collaboration with security teams. The ideal candidate has expertise in cloud-native technologies and scripting, contributing to impactful AI solutions. This is a full-time position at a mid-senior level.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - AI-Driven SaaS Platform
Site Reliability Engineer - AI-Driven SaaS Platform

Obsidian Security • Salford

On-site
GBP 85,000 - 103,000
Competitive compensation with equity
Comprehensive healthcare
Flexible paid time off
+2
Site Reliability Engineer - Cloud, Kubernetes & Observability
Site Reliability Engineer - Cloud, Kubernetes & Observability

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Kubernetes SRE: Cloud-Native Reliability & Automation
Kubernetes SRE: Cloud-Native Reliability & Automation

iXceed Solutions • Basildon

On-site
GBP 55,000 - 75,000
Site Reliability Engineer
Site Reliability Engineer

Helsing • City Of London

On-site
GBP 60,000 - 80,000
Lead Site Reliability Engineer: Low-Latency Linux & HPC
Lead Site Reliability Engineer: Low-Latency Linux & HPC

Autonomai Recruitment • England

On-site
GBP 70,000 - 90,000
Senior Site Reliability Engineer - Cloud & Automation
Senior Site Reliability Engineer - Cloud & Automation

ScaleneWorks People Solutions LLP • Bournemouth

On-site
GBP 60,000 - 80,000
Site Reliability Engineer (DV Security Clearance)
Site Reliability Engineer (DV Security Clearance)

Onyx-Conseil • Manchester

On-site
GBP 90,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

iXceed Solutions • Basildon

On-site
GBP 55,000 - 75,000
Site Reliability Engineer (SRE) - Cloud Kubernetes Platform
Site Reliability Engineer (SRE) - Cloud Kubernetes Platform

Intuition IT Solutions Ltd • Glasgow

Hybrid
GBP 60,000 - 75,000
Hybrid work model
Senior Site Reliability Engineer -Cloud,CI/CD & Kubernetes
Senior Site Reliability Engineer -Cloud,CI/CD & Kubernetes

Alchemy • Reading

Hybrid
GBP 60,000 - 80,000
Competitive salary
Healthcare benefits