Compute Platform Engineer — Multi-Cloud GPU & Kubernetes

Reflection AI

Greater London

On-site

GBP 80,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Top-tier compensation
Comprehensive medical, dental, and vision insurance
Fully paid parental leave
Paid time off
Daily lunch and dinner provided

Job summary

A technology company in the UK is seeking a skilled team member to enhance their Compute Platform. The role focuses on managing a K8s-based platform, ensuring system health, and improving performance. Responsibilities include cluster management, designing effective monitoring strategies, and preparing infrastructure for future GPU deployments. Ideal candidates will have strong systems-level engineering abilities, cloud storage expertise, and deep knowledge of GPU hardware in a Kubernetes environment. This position offers top-tier compensation and a collaborative work environment.

Qualifications

  • Experience focusing on cluster-wide behavior and maintenance.
  • Proven skills in systems or GPU infrastructure.
  • Familiarity with NCCL and multi-GPU environments.

Responsibilities

  • Build and maintain tools for automatic remediation and capacity planning.
  • Design cluster management stack for large-scale workloads.
  • Implement cluster-wide monitoring and performance benchmarking.
  • Prepare infrastructure for next-generation GPU deployments.

Skills

Systems-level engineering experience
Strong coding ability
Deep GPU hardware knowledge
Alignment with a K8s-first architecture
Cloud storage expertise

Job description

A technology company in the UK is seeking a skilled team member to enhance their Compute Platform. The role focuses on managing a K8s-based platform, ensuring system health, and improving performance. Responsibilities include cluster management, designing effective monitoring strategies, and preparing infrastructure for future GPU deployments. Ideal candidates will have strong systems-level engineering abilities, cloud storage expertise, and deep knowledge of GPU hardware in a Kubernetes environment. This position offers top-tier compensation and a collaborative work environment.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer — GPU HPC & Bare-Metal Clusters
Platform Engineer — GPU HPC & Bare-Metal Clusters

CATCHES • United Kingdom

Remote
GBP 50,000 - 70,000
Compute Platform Engineering Lead — Multi-Cloud & GPU
Compute Platform Engineering Lead — Multi-Cloud & GPU

Reflection • Greater London

On-site
GBP 120,000 - 180,000
Top-tier compensation
Stock options
Health & wellness
+5
Senior Platform Engineer (GCP) - Kubernetes & Cloud Infra
Senior Platform Engineer (GCP) - Kubernetes & Cloud Infra

Clavium • Greater London

Hybrid
GBP 70,000 - 90,000
Founding Cloud Platform Engineer – Multi‑Cloud & Kubernetes
Founding Cloud Platform Engineer – Multi‑Cloud & Kubernetes

WRITER • Greater London

On-site
GBP 120,000 - 180,000
Generous PTO, plus company holidays
Comprehensive medical and dental insurance
Paid parental leave for 12 weeks
+5
Compute Platform Lead — Senior Systems Architect (Multi-Cloud)
Compute Platform Lead — Senior Systems Architect (Multi-Cloud)

Reflection • Greater London

On-site
GBP 140,000 - 200,000
Stock options
Health & wellness
Meals in office
+5
HPC Network Engineer - GPU Cloud Infra (Remote UK)
HPC Network Engineer - GPU Cloud Infra (Remote UK)

asobbi • United Kingdom

Remote
GBP 53,000 - 69,000
Platform Engineer - AI Infra & Large-Scale GPU Systems
Platform Engineer - AI Infra & Large-Scale GPU Systems

Ineffable • Greater London

On-site
GBP 90,000 - 110,000
Senior Platform Engineer — DevOps, Kubernetes & Cloud
Senior Platform Engineer — DevOps, Kubernetes & Cloud

TripConnect • Greater London

Hybrid
GBP 70,000 - 90,000
Competitive compensation packages
Flexible work arrangements
Tuition assistance
+2
Compute Platform Engineer II: Scalable HPC & Cloud
Compute Platform Engineer II: Scalable HPC & Cloud

GSK • Greater London

On-site
GBP 60,000 - 80,000
Platform Engineer: GPU Inference & Training
Platform Engineer: GPU Inference & Training

Perplexity • Greater London

On-site
GBP 100,000 - 160,000