Site Reliability Engineer - Lisbon

Capital on Tap

Lisbon (IA)

Hybrid

USD 73,000 - 101,000

Full time

11 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Private Healthcare
Anniversary Rewards
Paid Car Parking
36 days holiday
Learning and Wellbeing Budget
Therapy sessions
Free drinks and snacks

Job summary

Capital On Tap in Lisbon is seeking a Site Reliability Engineer to join our hybrid team. You will design, build, and monitor scalable services to keep platforms fast, reliable, and observable.

You’ll automate cloud infrastructure (Azure), CI/CD pipelines (Azure DevOps, Octopus, Flux), and monitoring (DataDog); collaborate with Product and Platform teams to reduce toil and improve performance.

Qualifications

  • Experience in managing a public cloud (Azure advantageous).
  • Experience in Azure DevOps, Octopus, Flux or other CI/CD tools.
  • Experience with Linux and Microsoft Systems.
  • Excellent communication skills and ability to collaborate with multiple teams in an agile environment.
  • Proficient in contributing to IaC technologies involving expertise in writing, managing, and optimising infrastructure with tools such as Terraform and Pulumi.
  • Experience working with a cloud monitoring solution (advantageous to have DataDog).
  • Experience with Kubernetes and Docker.
  • Experience in at least one scripting language (Python, PowerShell, Go).

Responsibilities

  • Manage and automate Azure, Datadog, NGINX & Cloudflare.
  • Develop and monitor Kubernetes and Serverless resources.
  • Maintain infrastructure code with Terraform & CRDs / Crossplane.
  • Improve systems, processes, and technologies; consult stakeholders to enhance platform performance.
  • Getting involved in new application architecture & design processes.
  • Design solutions to reduce toil, automate repetitive tasks and streamline workflows to reduce manual work and boost team productivity.
  • Create SLIs and SLOs; increase application visibility.
  • Align with the Product team on SLAs and core service objectives.
  • Collaborate with Platform Engineers for automated solutions and pipelines.
  • Enhance user experience with infrastructure and pipeline optimisation.
  • Support CI/CD tools such as Azure Devops, Octopus Deploy and Flux to streamline software delivery.
  • Lead incident troubleshooting to safeguard customer experience

Skills

Cloud experience
Communication
Collaboration

Tools

Azure
Azure DevOps
Octopus Deploy
Flux
Linux
Terraform
Pulumi
Datadog
Kubernetes
Docker
Python
PowerShell
Go

Job description

We’re Capital on TapCapital on Tap started because small businesses were underserved. Big banks were slow, their products weren't fit for purpose, and small business owners often couldn't access what they needed. We set out to fix that.

Today we're a financial platform - not just a credit card company. We offer a best-in-class business credit card, SME-focused spend management platform, a savings product that hit £1 billion in funds within its first year, and a growing suite of tools and financial products that make running a small business easier.

1,000+ employees, £20bn in annual card spend, 200,000+ customers, 17,000+ Trustpilot reviews averaging 4.7 stars, and we're profitable. We’ve done a pretty good job so far, but we’re just getting started!

Lisbon | 2 Days in Office

SRE at Capital On Tap

At Capital On Tap, we run a hybrid embedded SRE model. We aim to work closely with the teams within Capital On Tap to provide them the best support. Our main objective currently is to gain as much visibility into our platform's health while offering scalable solutions.

What You’ll be doing:

As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You’ll design, build, and monitor systems, prevent issues before they happen. Using SLAs, SLIs, and SLOs, you’ll guide feature launches while maintaining services that everyone can depend on.

  • Manage and automate Azure, Datadog, NGINX & Cloudflare
  • Develop and monitor Kubernetes and Serverless resources
  • Maintain infrastructure code with Terraform & CRDs / Crossplane
  • Improve systems, processes, and technologies; consult stakeholders to enhance platform performance
  • Getting involved in new application architecture & design processes
  • Design solutions to reduce toil, automate repetitive tasks and streamline workflows to reduce manual work and boost team productivity.
  • Create SLIs and SLOs; increase application visibility
  • Align with the Product team on SLAs and core service objectives
  • Collaborate with Platform Engineers for automated solutions and pipelines
  • Enhance user experience with infrastructure and pipeline optimisation
  • Support CI/CD tools such as Azure Devops, Octopus Deploy and Flux to streamline software delivery
  • Lead incident troubleshooting to safeguard customer experience
Our Values & Culture
  • Just Pilot: We never settle for “good enough”. We pilot new ideas fast, ask questions to figure it out, and scale quickly.
  • Why Not Today? Fast is as slow as we go - speed and simplicity gives us a competitive advantage.
  • Be a Buddy: We tap in from day one to help the team, we do the right thing even if it’s hard.
  • Owners and Dates: We don’t chase people. If you own a task and agree to a date, the expectation is that it gets done.
  • Feedback: We want our employees to flourish, so we regularly provide direct and constructive feedback.
We’re Looking For

Required skills:

  • Experience in managing a public cloud (Azure advantageous)
  • Experience in Azure DevOps, Octopus, Flux or other CI/CD tools
  • Experience with Linux and Microsoft Systems
  • Excellent communication skills and ability to collaborate with multiple teams in an agile environment
  • Proficient in contributing to IaC technologies involving expertise in writing, managing, and optimising infrastructure with tools such as Terraform and Pulumi
  • Experience working with a cloud monitoring solution (advantageous to have DataDog)
  • Experience with Kubernetes and Docker
  • Experience in at least one scripting language (Python, PowerShell, Go)
Interview Process
  • First stage: 30-minute intro, CV review, and values with Talent Partner
  • Second stage: 75 minute Technical Task
  • Final stage: 30 minute Interview with EM + 30 minute Interview with Head of Platform Engineering
Diversity & Inclusion

We welcome, consider and encourage applications from anyone who shares our commitment to inclusivity. Join us in creating a space where authenticity thrives, and everyone can do their best work.

Great Work Deserves Great Perks

We try not to take ourselves too seriously (all the time) so we make sure our office is decked out with a pool table, arcade machine, beer tap, and a couple of office dogs thrown in for good measure.

Check out our benefits:

  • Private Healthcare through AdvanceCare
  • Anniversary Rewards (€280, €570, €860, 4-week fully paid sabbatical)
  • Paid Car Parking
  • 36 days holiday (inclusive of public holidays with public holidays carried over if they fall on a weekend)
  • Annual Learning and Wellbeing Budget
  • 6 free therapy sessions per year
  • Free drinks and snacks in our offices
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Customer Service Advisor - Midnights Team
Customer Service Advisor - Midnights Team

Remote Worker LTD. • United States

Remote
USD 34,000 - 41,000
Private Healthcare including dental/ey
Worldwide travel insurance
Anniversary Rewards
+10
Senior Database Engineer
Senior Database Engineer

Capital on Tap • London (KY)

Hybrid
USD 119,000 - 158,000
Private Healthcare
Travel insurance
Sabbatical
+9
Customer Service Advisor - Nights
Customer Service Advisor - Nights

Remote Worker LTD. • United States

Remote
USD 35,000 - 43,000
Private Healthcare including dental
Worldwide travel insurance
Anniversary Rewards
+10
Lead Automation Engineer
Lead Automation Engineer

Capital on Tap • Lisbon (IA)

Remote
USD 73,000 - 84,000
Private Healthcare
Anniversary Rewards
Paid Car Parking
+4
Software Engineer (Fully Remote)
Software Engineer (Fully Remote)

Capital on Tap • United States

Remote
USD 90,000 - 130,000
Private Healthcare
Generous holiday (36 days)
Learning & wellbeing budget
+1
Senior Software Engineer
Senior Software Engineer

Remote Jobs • United States

On-site
USD 120,000 - 170,000
Site Reliability Engineer
Site Reliability Engineer

sportygroup • United States

Remote
USD 150,000 - 210,000
Remote-first company
Competitive salary
Weekend Site Reliability Engineer
Weekend Site Reliability Engineer

Sporty Group • United States

On-site
USD 120,000 - 180,000
Remote first
Bonuses (quarterly)
28 days leave
+4
Weekend Site Reliability Engineer
Weekend Site Reliability Engineer

sportygroup • United States

Remote
USD 140,000 - 170,000
Remote-first company
Competitive salary with quarterlyBonu
28 days paid annual leave
+4
Team Lead Software Engineer
Team Lead Software Engineer

United States Digital Space LLC • United States

On-site
USD 73,000 - 104,000
Private Healthcare through AdvanceCare
Anniversary Rewards (€280, €570, €860)
36 days holiday
+3