DevOps Engineer

Encord

Greater London

Hybrid

GBP 90,000 - 120,000

Full time

28 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Competitive salary & equity
London office culture – in-person 4+ d
25 days annual leave
Learning & development budget
Travel for customer visits and events
Company lunches
Team socials & offsites

Job summary

Encord, the universal data layer for AI, is seeking a DevOps Engineer with 4+ years to join our platform engineering team in London. You’ll be embedded in squads building and operating Encord’s core infrastructure, ensuring performance, reliability, observability and scalability.

You’ll drive automation, design cloud infrastructure on GCP/AWS, manage Kubernetes, implement SLOs/SLAs, and champion IaC, monitoring, and incident response while partnering with multiple teams.

Qualifications

  • 4+ years of hands-on DevOps or SRE experience in production.
  • Experience building CI/CD pipelines and deployment automation at scale.
  • Proficiency with IaC tools and cloud platforms (GCP/AWS).

Responsibilities

  • Own and continuously improve deployment pipelines; collaborate with developers on infrastructure changes.
  • Design, deploy, and maintain cloud infrastructure on GCP and AWS; manage Kubernetes clusters and networking.
  • Drive automation and internal tooling to boost developer productivity and reduce manual toil.
  • Profile and optimize services for large-scale data pipelines; establish performance benchmarks.
  • Define and own SLIs/SLOs/SLAs; build alerting, runbooks, and incident response processes.
  • Instrument services with distributed tracing, logging, and metrics; ensure observability before production.

Skills

CI/CD pipelines
Kubernetes
IaC (Terraform/Pulumi)
Observability
Distributed systems
Networking
Databases
Cloud platforms (GCP/AWS)
SRE/Reliability engineering
Python

Tools

Prometheus
Grafana
OpenTelemetry
GCP
AWS
Docker
Kubernetes

Job description

About us

Encord is the universal data layer for AI that helps 300+ AI teams train and run models on the right data. Our platform indexes, curates, annotates, and evaluates data across the full AI lifecycle, from development through production. Trusted by Woven by Toyota, AXA, UiPath, Zipline, and more. We're an ambitious team of 100+ working at the frontier of AI and have raised $60M in Series C funding from Wellington Management, CRV, Next47 and Y Combinator.

The role

We're looking for a DevOps Engineer with 4+ years of experience to join our growing platform engineering team in London. You'll be embedded in the teams building and operating Encord's core infrastructure, ensuring our platform is performant, reliable, observable, and scalable. You'll drive a culture of automation, performance, and resilience through individual contributions and collaboration with multiple squads.

What You'll Do
  • CI/CD & Deployment Own and continuously improve deployment pipelines; partner closely with developers to review infrastructure changes, streamline release processes, and champion DevOps best practices across the engineering group.
  • Infrastructure & Cloud Design, deploy, and maintain cloud infrastructure on GCP and AWS; manage Kubernetes clusters, networking, and storage at petabyte scale, with infrastructure-as-code as the default.
  • Automation & Tooling Drive developer productivity by building, guiding, and reviewing automation and internal tooling efforts across the engineering group; eliminate manual toil wherever it exists.
  • Performance & Capacity Profile and optimise services handling large-scale data pipelines; perform capacity planning for storage and compute-intensive workloads. Work with squads to establish performance benchmarks and expectations.
  • Reliability & Availability Define and own SLIs/SLOs/SLAs for critical services; build alerting, runbooks, and incident response processes; lead postmortems with a blameless culture.
  • Observability Instrument services with distributed tracing, logging, and metrics (Prometheus, Grafana, OpenTelemetry, GCP Dashboards or similar); build infrastructure, define best practices, and work with each squad to ensure every service is observable before it goes to production.
What We're Looking For
  • 4+ years of hands-on DevOps, platform engineering, or SRE experience in a production environment.
  • Strong experience building and maintaining CI/CD pipelines and deployment automation at scale.
  • Proven experience with infrastructure-as-code tools (e.g., Terraform, Pulumi) and configuration management.
  • Strong fundamentals in designing, building, and maintaining resilient distributed and/or high performance systems.
  • Hands-on experience with Kubernetes and containerised workloads in cloud environments (GCP and/or AWS).
  • Solid understanding of networking, operating systems, and database technologies.
  • Experience with observability fundamentals metrics, logs, traces, and alerting.
Tech stack
  • Backend: Python
  • Frontend: TypeScript and React
  • Deployment: Kubernetes
  • Infrastructure: GCP
  • Machine learning: PyTorch, CUDA, Ray
Why Encord
  • Competitive salary, commission, and meaningful equity in a high-growth startup
  • Strong in-person culture - most of the team works from our London office 4+ days/week
  • 25 days annual leave + UK public holidays
  • Annual learning & development budget
  • Travel for customer visits, events, and conferences across the UK and Europe
  • Company lunches twice a week
  • Monthly socials & bi-annual team offsites
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Full-Stack
Senior Software Engineer, Full-Stack

The Consensus • Greater London

Hybrid
GBP 90,000 - 125,000
London office
Learning budget
25 days annual leave
+3
Senior Software Engineer - Backend
Senior Software Engineer - Backend

Visa Hunt • Greater London

On-site
GBP 90,000 - 130,000
Company equity
London office 4+ days/week
25 days leave
+4
Product Engineer
Product Engineer

Crane Venture Partners • Greater London

On-site
GBP 60,000 - 80,000
Competitive salary
Commission and equity
25 days annual leave
+3
Product Engineer
Product Engineer

The Consensus • Greater London

Hybrid
GBP 90,000 - 130,000
Competitive salary
Equity
London office 4+ days/week
+5
Forward Deployed Software Engineer
Forward Deployed Software Engineer

Crane Venture Partners • Greater London

On-site
GBP 70,000 - 110,000
London office 4+ days/week
Learning & development budget
25 days annual leave + UK holidays
+2
Solutions Engineer
Solutions Engineer

Crane Venture Partners • Greater London

Hybrid
GBP 90,000 - 130,000
Salary + equity
London office access
25 days leave
Software Engineer, Front-End Leaning
Software Engineer, Front-End Leaning

Crane Venture Partners • Greater London

On-site
GBP 60,000 - 80,000
Competitive salary
Equity in startup
25 days annual leave
+3
Forward Deployed Engineer
Forward Deployed Engineer

The Consensus • Greater London

Hybrid
GBP 90,000 - 130,000
Competitive salary
Exposure to cutting-edge AI
London office 4+ days/week
+2
Technical Program Manager
Technical Program Manager

Encord • Greater London

Hybrid
GBP 90,000 - 120,000
Equity in startup
London office 4+ days/week
25 days annual leave
+4
Technical Program Manager
Technical Program Manager

Crane Venture Partners • Greater London

Hybrid
GBP 70,000 - 110,000
Competitive salary & equity
London office culture
25 days leave + holidays
+4