Senior Infrastructure Engineer

Greenhouse Software, Inc.

Greater London

Hybrid

GBP 90,000 - 130,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Annual discretionary bonus
Private medical insurance
Flexible remote or hybrid work
Career progression
Collaborative international culture
Ownership and autonomy

Job summary

NexGen Cloud is the company behind Hyperstack, offering on-demand and private GPU infrastructure for researchers and enterprises. We’re a fast-moving team expanding OpenStack and Kubernetes globally, owning critical infrastructure that impacts performance, reliability, and customer experience.

This role focuses on ownership and end-to-end platform design and operations, not maintenance, with direct impact and rapid learning in a dynamic environment.

Qualifications

  • Exposure to GPU infrastructure, HPC, or large-scale compute environments.
  • Proven experience operating Kubernetes at scale, ideally bare-metal or private cloud.
  • Solid understanding of Linux, networking, and storage systems.
  • Experience with infrastructure automation, CI/CD, and Git-based workflows.
  • Strong ownership mindset with ability to operate without heavy oversight.

Responsibilities

  • Own the design, deployment, and operation of OpenStack and Kubernetes environments for GPU workloads.
  • Build and improve infrastructure using infrastructure-as-code and GitOps practices.
  • Optimise GPU workload scheduling and implement monitoring, logging, and alerting.
  • Lead incident response and drive reliability improvements across the platform.
  • Maintain strong security controls across infrastructure and container layers.

Skills

GPU infra
Kubernetes
OpenStack
Linux
Networking
Automation
GitOps
Incident response
Security controls

Tools

NVIDIA tooling

Job description

NexGen Cloud is the company behind Hyperstack, a full-stack AI cloud serving tens of thousands of customers from AI researchers to enterprises running the world's most compute-intensive workloads. We deliver on-demand and private GPU infrastructure to teams who treat performance as a requirement, not a feature.

We're a tight-knit, fast-moving team working at the cutting edge of AI cloud infrastructure. We practice what we preach, equipping our people with AI at every level so we can solve harder problems, ship faster, and keep raising the bar for what enterprise GPU infrastructure looks like.

This role exists because our platform is scaling quickly — and complexity comes with it. As we expand our OpenStack and Kubernetes environments globally, we need engineers who can take real ownership of how the platform is designed, operated, and improved. You'll have direct ownership over business-critical infrastructure that impacts performance, reliability, and customer experience.

This is not a maintenance role. If you like solving hard problems, owning systems end-to-end, and seeing the impact of your work immediately — you'll enjoy this.

WHAT YOU'LL BE DOING:

Rather than a long checklist, here's what success in this role looks like:

  • Own the design, deployment, and operation of OpenStack and Kubernetes environments — ensuring platform performance, scalability, and resilience for GPU workloads
  • Build and improve infrastructure using infrastructure-as-code and GitOps practices, driving automation across provisioning, deployment, and operational workflows
  • Optimise GPU workload scheduling using Kubernetes and NVIDIA tooling, and implement monitoring, logging, and alerting to ensure platform stability
  • Lead incident response and drive continuous improvement of reliability across the platform
  • Maintain strong security controls across infrastructure and container layers — RBAC, network policies, and tenant isolation
  • Work closely with Platform, DevOps, AI, Product, and Support teams to align infrastructure capabilities with customer and platform requirements
ABOUT YOU:

We're more interested in how you think and work than in a perfect CV. You'll likely bring a combination of the following:

  • Exposure to GPU infrastructure, HPC, or large-scale compute environments
  • Openstack experience in production and building openstack environments.
  • Proven experience operating Kubernetes at scale — ideally bare-metal or private cloud
  • Solid understanding of Linux, networking, and storage systems
  • Experience with infrastructure automation, CI/CD, and Git-based workflows
  • Strong ownership mindset — comfortable operating without heavy oversight and able to simplify and scale systems in a fast-moving environment
Nice to Have
  • Experience integrating Kubernetes with OpenStack
  • Ideally hands-on experience running OpenStack in production environments
  • Familiarity with advanced networking or cloud-native ecosystems
  • Contributions to open-source projects
WHAT WE OFFER:
  • Competitive salary and annual discretionary bonus scheme
  • 25 days of holiday, plus public holidays
  • Private medical insurance
  • Flexible working arrangements (remote or hybrid, depending on role and location)
  • Real ownership and autonomy, with the trust to take initiative and experiment
  • The opportunity to make a visible, meaningful impact as we scale
  • Clear career progression and growth opportunities in a fast-growing company
  • A collaborative, international culture built on trust, transparency, and ownership
  • The chance to help shape NexGen Cloud's team, culture, and future alongside ambitious, mission-driven colleagues
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Infrastructure Operations Engineer
Infrastructure Operations Engineer

NexGen Cloud • Greater London

Hybrid
GBP 75,000 - 110,000
Discretionary bonus
Flexible working
Wellbeing benefits
+1
Infra Operations Engineer – GPU Cloud & OpenStack
Infra Operations Engineer – GPU Cloud & OpenStack

NexGen Cloud • Greater London

Hybrid
GBP 75,000 - 110,000
Discretionary bonus
Flexible working
Wellbeing benefits
+1
Business Development Manager
Business Development Manager

NexGen Cloud • Greater London

On-site
GBP 80,000 - 110,000
Annual bonus
Wellbeing benefits
25 days holiday
+5
Strategic Sales Lead, AI Natives
Strategic Sales Lead, AI Natives

NexGen Cloud Ltd • Greater London

Hybrid
GBP 65,000 - 90,000
Competitive salary
Annual discretionary bonus
Flexible working arrangements
+1
Business Development Manager
Business Development Manager

NexGen Cloud Ltd • Greater London

Hybrid
GBP 70,000 - 90,000
Competitive salary and annual discretionary bonus
25 days of holiday plus public holidays
Flexible working arrangements
+1
Senior Infrastructure Engineer (Linux & Networking) - remote, US hours
Senior Infrastructure Engineer (Linux & Networking) - remote, US hours

NexGen Cloud • United Kingdom

Remote
GBP 60,000 - 80,000
100% home-office
Flexible working hours
Collaboration with a diverse team
+2
Lead GPU Infra Engineer — OpenStack & Kubernetes
Lead GPU Infra Engineer — OpenStack & Kubernetes

Greenhouse Software, Inc. • Greater London

Hybrid
GBP 90,000 - 130,000
Annual discretionary bonus
Private medical insurance
Flexible remote or hybrid work
+3
HPC Infrastructure Site Reliability Engineer
HPC Infrastructure Site Reliability Engineer

Radiant • Gloucester

On-site
GBP 90,000 - 120,000
Deployment Engineering Director, Systems Engineering
Deployment Engineering Director, Systems Engineering

Nscale • Greater London

On-site
GBP 120,000 - 180,000
Senior Business Development Manager - Vertical Markets
Senior Business Development Manager - Vertical Markets

Hyperstack • Greater London

Hybrid
GBP 90,000 - 140,000
Private Medical Insurance
25 days holiday
Hybrid working
+2