Platform Engineer

CATCHES

United Kingdom

Remote

GBP 50,000 - 70,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Co-working allowances
High-trust environment
Cutting-edge technology

Job summary

A luxury fashion technology company in the United Kingdom is seeking a Platform Engineer to manage and optimise high-performance computing infrastructure. The role involves automating provisioning of GPU clusters, maintaining Linux environments, and implementing monitoring solutions. Ideal candidates have a strong background in Linux systems, experience with Bare Metal servers, and proficiency in IaC tools. This full-time position is fully remote with a focus on innovation and low bureaucracy.

Qualifications

  • Strong background in Linux Systems Administration.
  • Experience managing Bare Metal servers (on-premise or packet/equinix metal).
  • Proficiency in Infrastructure as Code (IaC) tools.

Responsibilities

  • Automate the provisioning and lifecycle of high-performance GPU clusters.
  • Maintain the stability and performance of large-scale Linux environments.
  • Collaborate with vendors and internal teams to troubleshoot hardware and networking bottlenecks.
  • Implement monitoring solutions to visualise GPU health and cluster efficiency.
  • Assist in optimising the stack for containerised workloads.

Tools

Terraform
Ansible
Prometheus
Grafana
Kubernetes
Docker

Job description

Backed by some of the most influential names in luxury fashion globally. We blend advanced 3D rendering, AI and VFX techniques to deliver unparalleled shopping experiences for luxury fashion.

Role

We are hiring a Platform Engineer to manage and optimise our next-generation high-performance computing infrastructure. Move beyond standard cloud instances and manage the raw power of bare metal GPU clusters.

Responsibilities
  • Automate the provisioning and lifecycle of high-performance GPU clusters using Terraform and Ansible.
  • Maintain the stability and performance of large-scale Linux environments supporting AI/ML training workloads.
  • Collaborate with vendors and internal teams to troubleshoot hardware and networking bottlenecks (latency, throughput).
  • Implement monitoring solutions (Prometheus/Grafana) to visualise GPU health and cluster efficiency.
  • Assist in optimising the stack for containerised workloads (Kubernetes/Docker).
Requirements
  • Strong background in Linux Systems Administration.
  • Experience managing Bare Metal servers (on-premise or packet/equinix metal).
  • Proficiency in Infrastructure as Code (IaC) tools.
  • Nice to have: Exposure to GPUs, InfiniBand, or high-throughput networking (we will train the right candidate).
What working with CATCHES is like
  • Fully remote-first, async-friendly, with optional co-working allowances.
  • High-trust, low-bureaucracy environment that values experimentation and shipping.
  • Early influence on product, architecture and engineering culture.
  • Cutting-edge tech, luxury-fashion creativity, and games-industry scale challenges combined.
Seniority level
  • Mid-Senior level
Employment type
  • Full-time
Job function
  • Information Technology
Industries
  • Technology, Information and Internet and Retail Apparel and Fashion

Referrals increase your chances of interviewing at CATCHES by 2x

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer — GPU HPC & Bare-Metal Clusters
Platform Engineer — GPU HPC & Bare-Metal Clusters

CATCHES • United Kingdom

Remote
GBP 50,000 - 70,000
Senior Platform Engineer (Product Initiatives) - Systems Integrator
Senior Platform Engineer (Product Initiatives) - Systems Integrator

Hamilton Barnes Associates Limited • United Kingdom

Remote
GBP 120,000 - 190,000
High-Upside Equity
Flexible remote setup
Work-Life Balance
+1
GPU Infrastructure Lead - Systems Integrator
GPU Infrastructure Lead - Systems Integrator

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 140,000 - 170,000
Full Benefits
Network Engineer
Network Engineer

asobbi • United Kingdom

Remote
GBP 53,000 - 69,000
Highly competitive package with equity
Dynamic progression plan
Human-first flexibility
Performance Engineering Manager
Performance Engineering Manager

G-Research • Greater London

On-site
GBP 95,000 - 130,000
Highly competitive compensation
Lunch provided
35 days’ annual leave
+5
Machine Learning Performance Engineer
Machine Learning Performance Engineer

G-Research • Greater London

Hybrid
GBP 90,000 - 150,000
Competitive pay
Lunch provided
Annual leave 35d
+5
Platform Engineer - Financial Services
Platform Engineer - Financial Services

Hamilton Barnes Associates Limited • Greater London

Remote
GBP 135,000 - 165,000
Generous equity allocation
Remote Working (Travel to London)
Meaningful ownership over core platform architecture
+1
Platform Engineer
Platform Engineer

LinuxRecruit • Greater London

Hybrid
GBP 70,000 - 90,000
Platform Engineer
Platform Engineer

Carbon3ai Limited. • United Kingdom

Hybrid
GBP 90,000 - 120,000
Senior ML Infrastructure Engineer
Senior ML Infrastructure Engineer

Ellison Institute of Technology Oxford • Oxford

On-site
GBP 60,000 - 80,000
Enhanced holiday pay
Pension
Life Assurance
+6