Senior AI Systems Administrator

CommonAI C.I.C.

Cambridge

On-site

GBP 60,000 - 90,000

Full time

7 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Collaborative environment
Growth opportunities
Competitive salary + pension
Professional development
Networking opportunities
Office near Cambridge station

Job summary

CommonAI CIC, a non-profit membership organisation, is seeking a Senior AI Systems Administrator to provision and maintain multi-rack GPU server clusters for inference and training workloads in Cambridge. The role blends high-end hardware with software orchestration and collaborative engineering across teams.

You will work with cutting-edge AI tooling, optimise systems, and support users with scalable, robust infrastructure during UK business hours.

Qualifications

  • Significant experience as a System Administrator, HPC Engineer, MLOps Engineer or equivalent role.
  • Modern GPU server deployment, tuning and management.
  • Advanced networking technologies (Infiniband, RDMA, RoCE).
  • Advanced Linux skills with servers running hypervisors/VMs and multiple distros (Ubuntu, RHEL).
  • Expertise in source code management (Git) and infrastructure-as-code tools (Terraform, OpenTofu).

Skills

GPU deployment
Linux administration
Networking
Terraform
OpenTofu
Git
Hypervisors/VMs
Python scripting
Storage systems

Tools

Slurm
LSF
Ceph
Lustre
Weka
ROCE

Job description

CommonAI CIC is a non-profit membership organisation, founded on a belief in collaborative engineering for the safe and responsible development of foundational AI technologies. A place where AI startups, enterprises large and small, public sector bodies and academia can share resources and knowledge, to codevelop and grow businesses, fast.

We are led by experienced founders, investors and engineers who believe that collaborative engineering drives faster AI innovation and are backed by a mix of UK Government and private funding in order to design, build and deploy innovative AI systems.

The Opportunity

We are seeking a highly skilled Senior AI Systems Administrator to join our rapidly growing engineering team. You will provision and maintain multi-rack GPU server clusters designed for both inference and training workloads. This role is perfect for a hands-on engineer who thrives on the intersection of cutting-edge high-end hardware and complex software orchestration.

You will be given access to cutting edge AI tooling and we encourage our employees to make full use of the latest technology to find innovative and new ways of working.

You should be able to demonstrate:

  • Significant experience working as a System Administrator, HPC Engineer, MLOps Engineer or equivalent technical role
  • Modern GPU server deployment, tuning and management
  • Advanced networking technologies (Infiniband, RDMA, RoCE)
  • Advanced Linux skills and experience working with servers running hypervisors/VMs and multiple distributions (e.g. Ubuntu, RHEL)
  • Expertise in source code management (e.g. Git) and infrastructure-as-code tools (e.g. Terraform, OpenTofu)
  • A strong understanding of network design and switch/router/firewall configuration
  • The desire to work collaboratively with users and other stakeholders to iteratively optimise systems


Experience with any or all of the following will also be highly valued:

  • High-performance or high-availability storage servers/clusters (Lustre, Ceph, Weka, PEAK:AIO)
  • HPC workload managers (Slurm, LSF)
  • Proficiency in scripting (e.g. Python)

Our infrastructure is used for research and development. Support will generally only be required during UK business hours however major maintenance may occasionally be scheduled for weekends.

  • A collaborative and supportive work environment
  • The opportunity to have a high impact in a growing organisation
  • Competitive salary package, including pension and benefits
  • Professional development opportunities
  • Networking opportunities with influential figures from across the tech sector and academia
  • A vibrant office environment located a few minutes' walk away from Cambridge train station

CommonAI CIC is an equal opportunity employer and is committed to creating an inclusive and diverse workplace.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Systems Administrator
Senior AI Systems Administrator

CommonAI Holdings Ltd • Cambridge

Remote
GBP 65,000 - 90,000
Pension
Professional development
Networking opportunities
+1
Senior AI Systems & GPU HPC Engineer
Senior AI Systems & GPU HPC Engineer

CommonAI C.I.C. • Cambridge

On-site
GBP 60,000 - 90,000
Collaborative environment
Growth opportunities
Competitive salary + pension
+3
Senior AI Systems Administrator: GPU HPC Infra Lead
Senior AI Systems Administrator: GPU HPC Infra Lead

CommonAI Holdings Ltd • Cambridge

Remote
GBP 65,000 - 90,000
Pension
Professional development
Networking opportunities
+1
Senior Storage Architect
Senior Storage Architect

CommonAI CIC • Cambridge

On-site
GBP 60,000 - 75,000
Competitive salary package
Pension
Professional development opportunities
+2
Senior Performance Engineer
Senior Performance Engineer

CommonAI CIC • Cambridge

On-site
GBP 70,000 - 100,000
Competitive salary
Pension
Professional development
+2
Senior Software Engineer - AI-Native Cloud Infrastructure Cambridge, United Kingdom
Senior Software Engineer - AI-Native Cloud Infrastructure Cambridge, United Kingdom

CommonAI Compute Ltd. • United Kingdom

Remote
GBP 90,000 - 130,000
Stock options
Cambridge office
On-site gym
+1
Senior HPC Systems Administrator
Senior HPC Systems Administrator

1000scholars • Oxford

On-site
GBP 49,000 - 55,000
Senior Software Engineer - AI-Native Cloud Infrastructure
Senior Software Engineer - AI-Native Cloud Infrastructure

CommonAI CIC • Cambridge

On-site
GBP 75,000 - 110,000
Stock options
Competitive salary
Professional development
+2
Senior ML Infrastructure Engineer Enterprise Operations Oxford, England, United Kingdom
Senior ML Infrastructure Engineer Enterprise Operations Oxford, England, United Kingdom

Ellison Institute, LLC • Oxford

On-site
GBP 90,000 - 140,000
Competitive salary
25 days annual leave + 8 bank holidays
3 additional days between Christmas &.
+11
Senior Software Engineer (vLLM)
Senior Software Engineer (vLLM)

CommonAI Holdings Ltd • Cambridge

On-site
GBP 90,000 - 130,000
Competitive salary
Pension
Professional development
+2