Senior Site Reliability Engineer (SRE)

Nebius

Greater London

On-site

GBP 90,000 - 120,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Career growth
Learning opportunities
Flexibility
Ownership
Collaborative culture
Impactful AI projects
International environment

Job summary

Nebius is building a full‑stack AI cloud platform and continuously scales its services in a fast‑moving environment. You will join a team delivering high‑reliability infrastructure for data and model deployment, with emphasis on performance and availability.

The role focuses on building scalable backend systems, improving deployment pipelines, and applying cloud expertise to support cutting‑edge AI workloads across regions and platforms.

Qualifications

  • Proficient in Go, Python, or C++ with strong CS fundamentals.
  • Solid understanding of classic algorithms and data structures.
  • Commercial Unix and networking experience.
  • Experience with containerization and CM tools (Ansible, Salt, Terraform, Docker, Kubernetes, Helm).

Responsibilities

  • Ensure fault‑tolerance, scale, and uninterrupted operations for the service.
  • Use cutting‑edge cloud technology to solve a variety of infrastructure problems.
  • Implement and improve CI/CD processes.

Skills

Go
Python
C++

Tools

Docker
Kubernetes
Terraform
Ansible
Salt
Helm

Job description

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full‑stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in‑house AI/ML infrastructure. Built by engineers, for engineers. From large‑scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.

The Role
Your Responsibilities Will Include
  • Ensure fault‑tolerance, scale, and uninterrupted operations for the service.
  • Use cutting‑edge cloud technology to solve a variety of infrastructure problems.
  • Implement and improve CI/CD processes.
We Expect You To Have
  • Solid experience with programming languages like Go, Python, or C++.
  • Solid understanding of classic algorithms and data structures.
  • Commercial experience with and deep understanding of Unix systems and network technology.
  • Experience with systems for containerization and configuration management (Ansible, Salt, Terraform, Docker, K8s, Helm).
It Will Be An Added Bonus If You Have
  • Desire to be involved in backend development.
  • Experience designing, developing, and running high‑load distributed systems.
  • Commercial experience with a variety of cloud platforms.

We conduct coding interviews as part of the process.

Benefits & Perks
  • Competitive compensation
  • Career growth and learning opportunities
  • Flexibility and ownership
  • Collaborative and innovative cultureOpportunity to work on impactful AI projects
  • International environment and talented teams
What's It Like To Work At Nebius

Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI

Equal Opportunity Statement

Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (DevTools)
Senior Site Reliability Engineer (DevTools)

Nebius • Greater London

On-site
GBP 90,000 - 120,000
Competitive pay
Career growth
Flexible work
+3
Senior Software Engineer (Serverless)
Senior Software Engineer (Serverless)

Nebius • Greater London

Hybrid
GBP 120,000 - 170,000
Competitive compensation
Career growth and learning
Flexibility and ownership
+2
Senior Site Reliability Engineer — Token Factory (Inference Platform)
Senior Site Reliability Engineer — Token Factory (Inference Platform)

Nebius • Greater London

On-site
GBP 100,000 - 140,000
Competitive compensation
Career growth and learning opportunity
Flexibility and ownership
+3
Field Network Engineer
Field Network Engineer

Nebius • Greater London

On-site
GBP 42,000 - 65,000
Competitive pay
Career growth
Flexible work
+3
Data Center IT Manager
Data Center IT Manager

Nebius • Greater London

On-site
GBP 70,000 - 100,000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3
Data Center IT Technician
Data Center IT Technician

Nebius • Newport

On-site
GBP 28,000 - 36,000
Competitive pay
Career growth
Flexibility and ownership
+3
Senior Technical Project Manager (Region Delivery)
Senior Technical Project Manager (Region Delivery)

DeepCamp • Greater London

Hybrid
GBP 75,000 - 115,000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3
Senior Data Center IT Technician
Senior Data Center IT Technician

Nebius • Newport

On-site
GBP 30,000 - 42,000
Competitive compensation
Career growth and learning opportunity
Flexibility and ownership
+3
Senior Backend Developer (Token Factory)
Senior Backend Developer (Token Factory)

Nebius • Greater London

On-site
GBP 110,000 - 160,000
Competitive pay
Career growth
Flexibility
+3
Technical Program Manager - Compute Systems Engineering
Technical Program Manager - Compute Systems Engineering

Nebius • Greater London

On-site
GBP 70,000 - 110,000
Competitive compensation
Career growth and learning机会
Flexibility and ownership
+3