Site Reliability Engineer - Data

Capitalontap

Greater London

On-site

GBP 70,000 - 110,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Private Healthcare
Worldwide travel insurance
Sabbatical rewards
Pension scheme

Job summary

Capital On Tap is seeking a Site Reliability Engineer to join our hybrid SRE model. You will design, build, and monitor scalable data and platform systems embedded in the Data teams.

You will work with cloud, IaC, CI/CD, and monitoring tools to improve reliability and performance, collaborating with product and base engineering teams across Azure, Kubernetes, and SQL/MongoDB.

Qualifications

  • Experience managing public cloud environments.
  • Proficient in IaC with Terraform and related tools.
  • Experience with CI/CD pipelines and troubleshooting.
  • Proficient with Kubernetes and Docker.
  • Experience with a cloud monitoring solution.
  • Proficiency in scripting languages (Python, PowerShell, Go).
  • Experience with Data Warehouses or similar knowledge.
  • Experience with Snowflake or other DBMS platforms.
  • Strong SQL fundamentals and data architecture understanding.

Responsibilities

  • Manage and automate resources in Azure, Datadog, NGINX & Cloudflare.
  • Develop, deploy and monitor Kubernetes and Serverless resources.
  • Build, manage, and evolve IaC using Terraform, Helm and Go CRDs.
  • Improve systems and processes; collaborate with stakeholders to boost performance.
  • Design and maintain monitoring and alerting strategies with Datadog.
  • Contribute to new application architecture and design processes.
  • Optimize CI/CD pipelines and developer workflows with Azure DevOps, GitHub, Octopus Deploy, MirrorD, Flux.

Skills

Public cloud experience
IaC (Terraform)
CI/CD experience
Kubernetes & Docker
Scripting languages (Python/PowerShell
Data Warehouses/SQL

Tools

Terraform
Helm
Go CRDs
Datadog
NGINX
Cloudflare
Kubernetes
Docker
Azure DevOps
GitHub
Octopus Deploy
MirrorD
Flux

Job description

We’re Capital on Tap

Capital on Tap started because small businesses were underserved. Big banks were slow, their products weren’t fit for purpose, and small business owners often couldn’t access what they needed. We set out to fix that.

Today we're a financial platform - not just a credit card company. We offer a best-in-class business credit card, SME-focused spend management platform, a savings product that hit £1 billion in funds within its first year, and a growing suite of tools and financial products that make running a small business easier.

1,000+ employees, £20bn in annual card spend, 200,000+ customers, 17,000+ Trustpilot reviews averaging 4.7 stars, and we're profitable. We’ve done a pretty good job so far, but we’re just getting started!

London, Old Street | 2 Days in Office

SRE at Capital On Tap

At Capital On Tap, we run a hybrid embedded SRE model - We aim to work closely with the teams to provide them the best support. As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You’ll design, build, and monitor systems, preventing issues before they happen.

You will be embedded in the Data teams, being Data Platform who works alongside business teams, assisting with the ingestion and exports, model deployments and data warehouse management, and Database Engineering who are responsible for the maintenance and enhancement of the OLTP platform, based around SQL Server and MongoDB clusters, whilst also assisting software developers with data architecture, data transformation and performance tuning.

What you’ll be doing
  • Manage and automate resources in Azure, Datadog, NGINX & Cloudflare.
  • Develop, deploy and monitor Kubernetes and Serverless resources.
  • Build, manage, and evolve IAC using Terraform, Helm and Go CRDs.
  • Improving systems, processes, and technologies; consulting with stakeholders to enhance platform performance.
  • Design and maintain comprehensive monitoring and alerting strategies using Datadog.
  • Getting involved in new application architecture & design processes.
  • Designing solutions to reduce toil, automate repetitive tasks and streamline workflows to reduce manual work and boost team productivity.
  • Creating SLIs and SLOs; increasing application visibility
  • Align with the Product team on SLAs and core service objectives
  • Collaborating with core foundational teams to build and maintain reusable, automated solutions.
  • Optimising CI/CD pipelines and developer workflows using Azure DevOps, Github, Octopus Deploy, MirrorD, and Flux.
  • Leading incident response; taking an active part in communications, investigations, remediation and post-mortems to protect customer experience.
  • Help design and build systems that improve the reliability, resiliency and maintainability of our data systems and products.
  • Provide assistance and support to other teams focused on database related applications methodologies and system resources.
Our Values & Culture
  • Just Pilot: We never settle for “good enough”. We pilot new ideas fast, ask questions to figure it out, and scale quickly.
  • Why Not Today? Fast is as slow as we go - speed and simplicity gives us a competitive advantage.
  • Be a Buddy: We tap in from day one to help the team, we do the right thing even if it’s hard.
  • Owners and Dates: We don’t chase people. If you own a task and agree to a date, the expectation is that it gets done.
  • Feedback: We want our employees to flourish, so we regularly provide direct and constructive feedback.
What are we looking for
  • Experience managing public cloud environments.
  • Proficient in contributing to IaC technologies involving expertise in writing, managing, and optimising infrastructure with tools such as Terraform.
  • Experience using CI/CD tools, building pipelines, templates and troubleshooting.
  • Proficient with containerisation technologies such as Kubernetes and Docker.
  • Experience working with a cloud monitoring solution.
  • Proficiency in at least one scripting language such as Python, PowerShell, Go.
  • Great communication skills with the ability to collaborate effectively.
  • Experience with Data Warehouses or similar.
  • Experience with Snowflake or other DBMS platforms.
  • Expert in database fundamentals and SQL.

Even if you don’t have all of the necessary skills, we still encourage you to apply.

Interview process
  • First stage: 30 minute intro and values call with Talent Partner
  • Second stage: 60 minute CV overview and technical chat with the SRE Team Lead and the SRE & Platform Engineering Manager
  • Third stage: 75 minute technical exercise & questions with the SRE Lead
  • Final stage: 30 minute chat with the Lead DBA & Lead Data Platform Engineer
Diversity & Inclusion

We welcome, consider and encourage applications from anyone who shares our commitment to inclusivity. Join us in creating a space where authenticity thrives, and everyone can do their best work.

Great Work Deserves Great Perks

We try not to take ourselves too seriously (all the time) so we make sure our office is decked out with a pool table, arcade machine, beer tap, and a couple of office dogs thrown in for good measure. Check out our benefits:

  • Private Healthcare including dental and opticians services through Vitality
  • Worldwide travel insurance through Vitality
  • Anniversary Rewards (£250, £500, £750, 4-week fully paid sabbatical)
  • Salary Sacrifice Pension Scheme up to 7% match
  • 28 days holiday (plus bank holidays)
  • Annual Learning and Wellbeing Budget
  • Enhanced Parental Leave
  • Cycle to Work Scheme
  • Season Ticket Loan
  • 6 free therapy sessions per year
  • Dog Friendly Offices
  • Free drinks and snacks in our offices

Check out more of our benefits, values and mission here.

Other Info

Check out our Top Tips for interviewing.

Keep updated on new job opportunities by following us on Linkedin.

Email careers@capitalontap.com if you have any questions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Data
Site Reliability Engineer - Data

Capital on Tap • Greater London

Hybrid
GBP 90,000 - 120,000
Private Healthcare including dental
Worldwide travel insurance
Anniversary Rewards
+9
Site Reliability Engineer - Frontend
Site Reliability Engineer - Frontend

Capital on Tap • Greater London

Hybrid
GBP 70,000 - 120,000
Private Healthcare including dental &_
Worldwide travel insurance
Anniversary Rewards (£250, £500, £750)
+9
Cloud Engineering Team Lead
Cloud Engineering Team Lead

Capital on Tap • Greater London

On-site
GBP 90,000 - 120,000
Private healthcare including dental/ey
Annual learning budget
Dog-friendly offices
+3
Data Engineer
Data Engineer

Capital on Tap • Greater London

On-site
GBP 70,000 - 110,000
Healthcare
Travel Insurance
Anniversary Rewards
+10
People Director
People Director

Capital on Tap • Greater London

Hybrid
GBP 120,000 - 180,000
Private Healthcare including dental &
Worldwide travel insurance through Vit
Access to a women's health platform
+10
Senior Finance Data Scientist
Senior Finance Data Scientist

Capital on Tap • Greater London

Hybrid
GBP 90,000 - 120,000
Private Healthcare including dental
Worldwide travel insurance
Learning and wellbeing budget
+6
Mobile and Backend Software Engineer - Banking
Mobile and Backend Software Engineer - Banking

Capital on Tap • Greater London

Hybrid
GBP 90,000 - 130,000
Private Healthcare including dental
Worldwide travel insurance
Annual Learning and Wellbeing Budget
+4
Engineering Lead
Engineering Lead

Capital on Tap • Greater London

On-site
GBP 70,000 - 90,000
Private Healthcare including dental and opticians
Worldwide travel insurance
Anniversary Rewards
+9
People Director
People Director

Capitalontap • Greater London

On-site
GBP 120,000 - 180,000
Private Healthcare
Travel Insurance
Hertility access
+9
Senior Finance Data Scientist
Senior Finance Data Scientist

Capitalontap • Greater London

Hybrid
GBP 85,000 - 120,000
Private Healthcare including dental
Worldwide travel insurance
Cycle to Work Scheme
+4