Site Reliability Engineer

Genesis AI

Greater London

Hybrid

GBP 90,000 - 150,000

Full time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Genesis AI is seeking a Site Reliability Engineer to tackle reliability, scalability, and efficiency challenges across SRE and development teams. You will build and run large-scale distributed systems that keep Genesis platform reliable and performant for customers, and drive solutions across teams.

You will optimize existing systems, automate toil away, and guide technical decisions balancing system health with fast-moving product priorities.

Qualifications

  • Design, analyze and troubleshoot distributed systems at scale.
  • Experience with cloud platforms and container orchestration.
  • Leading complex, large-scale technical projects across teams.

Responsibilities

  • Take on ambiguous reliability, scalability, and efficiency challenges and drive solutions across SRE and development teams.
  • Build and run large-scale, massively distributed, fault-tolerant systems that keep Genesis platform reliable and performant for our customers.
  • Optimize existing systems, build infrastructure, and eliminate toil through automation to continuously improve uptime and rate of change.
  • Cultivate a culture of reliability throughout the organization, guiding technical decisions that balance system health with fast-moving product priorities.
  • Ensure the long-term health, maintainability, and reliability of services through capacity planning, performance analysis, and proactive incident prevention.

Skills

Python
Go
Distributed systems
Kubernetes
NALSD
Leadership

Tools

Kubernetes
Cloud Functions

Job description

What You'll Do
  • Take on ambiguous reliability, scalability, and efficiency challenges and drive solutions across SRE and development teams.

  • Build and run large-scale, massively distributed, fault-tolerant systems that keep Genesis platform reliable and performant for our customers.

  • Optimize existing systems, build infrastructure, and eliminate toil through automation to continuously improve uptime and rate of change.

  • Cultivate a culture of reliability throughout the organization, guiding technical decisions that balance system health with fast-moving product priorities.

  • Ensure the long-term health, maintainability, and reliability of services through capacity planning, performance analysis, and proactive incident prevention.

What You'll Bring
  • Strong software engineering skills (e.g., in Python, Go, or similar) with extensive experience designing, analyzing, and troubleshooting distributed systems.

  • Deep expertise with cloud computing platforms (e.g., Kubernetes, Cloud Functions) and Non-Abstract Large Systems Design (NALSD).

  • Experience leading complex, large-scale technical projects and providing technical leadership across teams.

  • Ability to apply coding, algorithms, and complexity analysis to solve ambiguous problems at scale with minimal disruption.

  • A collaborative, intellectually curious mindset - comfortable working across a wide variety of backgrounds and bringing cross-team perspective to build robust, reusable solutions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Greater London

Hybrid
GBP 90,000 - 130,000
SRE: Scale, Reliability & Automation Leader
SRE: Scale, Reliability & Automation Leader

Genesis AI • Greater London

Hybrid
GBP 90,000 - 150,000
Senior SRE: Scale, Reliability & Automation Leader
Senior SRE: Scale, Reliability & Automation Leader

Jobtailor • Greater London

Hybrid
GBP 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Site Reliability Engineer
Site Reliability Engineer

United States Digital Space LLC • Greater London

On-site
GBP 90,000 - 130,000
Daily catered lunches
Modern office environment
Tech talks and knowledge sharing
Head of Site Reliability Engineering
Head of Site Reliability Engineering

Goldman Sachs • Birmingham

On-site
GBP 60,000 - 100,000
Vice President - Site Reliability Engineering
Vice President - Site Reliability Engineering

Goldman Sachs • Birmingham

On-site
GBP 60,000 - 100,000
Site Reliability Engineer
Site Reliability Engineer

iXceed Solutions • Basildon

On-site
GBP 55,000 - 75,000
SRE Engineer
SRE Engineer

Savant Recruitment • Greater London

On-site
GBP 60,000 - 80,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

P2P • Greater London

On-site
GBP 90,000 - 130,000