Senior Data Reliability Engineer

Elliptic Enterprises Limited

Greater London

On-site

GBP 90,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid working option
Remote working budget
Learning & Development budget
25 days annual leave
Private Health Insurance
Life Assurance

Job summary

A leading technology company in the UK is seeking a Senior DRE to enhance engagement with Site Reliability Engineering across teams. The candidate will drive initiatives that ensure the quality and reliability of large datasets and lead the on-call process. This role demands quick decision-making and the ability to inspire teams. Key benefits include hybrid working arrangements, a remote working budget, generous leave provisions, and a comprehensive health plan.

Qualifications

  • Proven experience enhancing quality and reliability of large datasets.
  • Experience leading site reliability for a high-volume SaaS product.
  • Supported distributed systems in AWS.

Responsibilities

  • Drive engagement with Site Reliability across engineering.
  • Lead on building out a framework for data quality.
  • Command incidents and ensure follow-up actions are completed.

Skills

Site Reliability Engineering
Data Quality
On-call Process Management
Distributed Systems
AWS
Kubernetes

Job description

Overview

This job is brought to you by Jobs/Redefined, the UK's leading over-50s age inclusive jobs board.

The impact you will have:

As Senior DRE, you will drive engagement with Site Reliability across the full breadth of engineering. You will hold every engineer and every team accountable in building highly-resilient, robust, reliable software. You will be part of a cross-functional, cross-discipline team of SMEs and on-callers, whose mission it is to keep our platform highly performant 24/7/365.

Responsible for a diverse suite of products, you will oversee SR of enterprise grade applications that sit on the critical path running 1000s of QPS. Elliptic is known for its extensive and reliable datasets and you will play a critical role in defining and building out a market-leading foundation for data quality and control. This means building the processes, culture, and frameworks that will power observability, quality, data lineage, and remediation to form an essential pillar of our data & intelligence platform.

What you will do

This is a cross team role, and you will have the full support of leadership and engineering in carrying out your responsibilities - it's not all down to you, but you will show the rest of us what good looks like.

  • Evangelise SRE & DRE across engineering
  • Lead the charge on building out a framework for data quality that will provide our customers with strong guarantees about the fidelity of our data as well support our marketing and revenue functions
  • SRE as a function define and own the on-call process:
    • Quickly establishing a strong working knowledge of our systems
    • Commanding incidents
    • Running mop-ups
    • Ensuring follow-up actions are completed to your schedule
    • Evaluating and improving our existing E2E on-call process
  • Take part in the on-call rotation, one week every 4-5 weeks (24x7x365 coverage)
  • Evaluate, manage and maintain our existing solutions for monitoring, alerting, paging, response, documentation
  • Report on uptime, availability, performance, etc across our product suite
  • Write post-mortems for both internal and external consumption
  • Represent our SRE & DRE function on sales calls with tier one enterprise financial institutions
  • Work with product, sales and customer service to define SLAs for different products and use cases
  • Work with internal product teams to define SLOs for internal consumption and measurement
  • Work with our engineering teams directly to embed DRE practices
You will be a great fit here if you:
  • Thrive under high pressure situations, and are able to make tough decisions quickly
  • Fail fast, own the failure; encourage a blame free engineering culture
  • Are an inspiring thought leader, and are able to take others with you on a journey
  • Aren't afraid to get your hands dirty and dig into code across myriad technologies
  • Understand the importance of reliability in enterprise finance systems
  • Have strong opinions based on your experience that you evolve over time as you learn from others
Our ideal candidate has:
  • Proven experience at leveling up the quality and reliability of large datasets not just services and APIs
  • Experience leading site reliability for a high volume SaaS product
  • Supported distributed systems in AWS
  • The presence and empathy required to hold teams to account
  • Defined SLAs / SLOs both internal and client facing
  • Offered post mortems to enterprise clients (verbal and written)
Bonus Points for:
  • Having a genuine interest in the crypto ecosystem and being behind the mission of the company
  • Working knowledge of Kubernetes and the challenges presented
Job Benefits
How we work:
  • Hybrid working and the option to work from almost anywhere for up to 90 days per year
  • £500 Remote working budget to set up your home office space
Learning & Development:
  • $1,000 Learning & Development budget to use on anything (agreed with your manager) that contributes to your growth and development
Vacation/ Leave:
  • Holidays: 25 days of annual leave + bank holidays
  • An extra day for your birthday
  • Enhanced parental leave: we provide eligible employees, regardless of gender or whether they become a parent by birth or adoption, 16 weeks fully-paid leave and leave.
Benefits:
  • Private Health Insurance - we use Vitality!
  • Full access to Spill Mental Health Support
  • Life Assurance: we hope you will never need this - but our cover is for 4 times your salary to your beneficiaries
  • £100 Crypto for you!
  • Cycle to Work Scheme
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps Engineer
Senior DevOps Engineer

Elliptic Enterprises Limited • Greater London

Hybrid
GBP 75,000 - 100,000
Hybrid working options
Remote working budget
Learning & Development budget
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LSEG • Nottingham

On-site
GBP 70,000 - 90,000
Healthcare
Retirement planning
Paid volunteering days
+1
Director of Site Reliability Engineering
Director of Site Reliability Engineering

EPAM Systems • Greater London

Hybrid
GBP 140,000 - 200,000
ESPP
Life assurance
Income protection
+11
Site Reliability Engineer - Data
Site Reliability Engineer - Data

Capitalontap • Greater London

Hybrid
GBP 70,000 - 110,000
Private Healthcare
Worldwide travel insurance
Sabbatical rewards
+1
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hitachids • Greater London

On-site
GBP 90,000 - 140,000
Senior SRE
Senior SRE

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
Engineering Manager
Engineering Manager

Elliptic Enterprises Limited • Greater London

Hybrid
GBP 75,000 - 95,000
Hybrid working options
£500 Remote working budget
$1,000 Learning & Development budget
+3
Site Reliability Engineer - Frontend
Site Reliability Engineer - Frontend

Capital on Tap • Greater London

Hybrid
GBP 70,000 - 120,000
Private Healthcare including dental &_
Worldwide travel insurance
Anniversary Rewards (£250, £500, £750)
+9
Lead Site Reliability Engineer
Lead Site Reliability Engineer

LSEG • Nottingham

On-site
GBP 80,000 - 100,000
Healthcare
Retirement Planning
Paid Volunteering Days
+1
Head of Site Reliability Engineering (SRE)
Head of Site Reliability Engineering (SRE)

United States Digital Space LLC • Greater London

Hybrid
GBP 150,000 - 210,000
ClassPass
Unlimited vacation
Apple equipment
+3