Senior Site Reliability Engineer

Fountain

Lisboa

Presencial

EUR 70 000 - 110 000

Tempo integral

Há 3 dias
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Uma candidatura completa num minuto — currículo e carta de apresentação personalizados, prontos a enviar.

Ultrapassa os filtros ATS

Resumo da oferta

Fountain is seeking a Site Reliability Engineer to own reliability, scalability, and operational excellence of its platform. You will partner with Engineering, Security and Product to build resilient infrastructure, improve developer experience, and raise the bar on observability and incident response.

You will design and maintain Kubernetes-based infrastructure, manage the AWS footprint, and advance CD/GitOps practices with ArgoCD.

Qualificações

  • 5+ years of site reliability engineering experience.
  • Experience owning reliability for production systems - defining SLOs or error budgets.
  • Hands-on Kubernetes/Helm on EKS in production.
  • Experience with Cloudflare CDNs/WAFs.
  • CI/CD or GitOps deployment experience.
  • Led or significantly contributed to incident response and postmortems.
  • Working knowledge of AWS core services beyond EKS.

Responsabilidades

  • Own platform reliability, SLOs/SLIs, capacity planning and production readiness.
  • Design, build and maintain Kubernetes-based infrastructure and deployment workflows.
  • Operate and evolve AWS footprint with a security-first mindset.
  • Improve CD/GitOps practices (ArgoCD) and deployment safety.
  • Build autoscaling strategies for services and workloads (KEDA).
  • Lead incident response: on-call, triage, mitigation, postmortems, preventative follow-ups.
  • Enhance observability with metrics, logs, traces and alerting (OpenTelemetry).
  • Collaborate with application teams to reduce toil and improve operational maturity.
  • Improve infrastructure-as-code practices and Terraform modules.

Ferramentas

Kubernetes
Helm
EKS
Cloudflare
CI/CD
GitOps
ArgoCD
Terraform
OpenTelemetry

Descrição da oferta de emprego

When you join the Fountain team, you become part of the leading enterprise solution for frontline workforce management. Fountain's automated, customizable platform provides a seamless applicant experience for workers, while ensuring organizations can scale and manage their frontline talent.

We've helped hundreds of companies like UPS, CLEAR, Stitch Fix, GoPuff, Fetch, and sweetgreen to hire, onboard, and manage over 14 million workers in more than 75 countries.

In 2022, we closed $185M in our Series C, led by SoftBank and B Capital.

Join our growing team of highly collaborative, ambitious, and forward-thinking Fountaineers as we empower our hundreds of customers and millions of frontline workers around the world.

Let's elevate frontline work together.

As a Site Reliability Engineer (SRE) at Fountain, you'll own the reliability, scalability, and operational excellence of the systems that power our platform. You'll partner closely with Engineering, Security, and Product to build resilient infrastructure, improve developer experience, and raise the bar on observability and incident response.

What you'll be doing:
  • Own and improve platform reliability (SLOs/SLIs), capacity planning, and production readiness
  • Design, build, and maintain Kubernetes-based infrastructure and deployment workflows
  • Operate and evolve our AWS footprint (networking, compute, storage, IAM) with a security-first mindset
  • Improve our CD/GitOps practices (ArgoCD) and deployment safety (progressive delivery, rollbacks, guardrails)
  • Build autoscaling strategies for services and workloads (KEDA where appropriate)
  • Lead incident response: on-call, triage, mitigation, postmortems, and preventative follow-through
  • Strengthen observability across services: metrics, logs, traces, and alerting (OpenTelemetry + dashboards)
  • Partner with application teams to tune performance, reduce toil, and improve operational maturity
  • Improve infrastructure-as-code practices and maintain Terraform modules and environments
  • Contribute to evaluating and integrating AI tooling and MCP tools to accelerate operational workflows
What you should bring:
  • 5+ years of experience in site reliability engineering
  • Experience owning reliability for production systems - defining SLOs or running error budgets
  • Deep hands-on experience with Kubernetes/Helm on EKS in production
  • Experience with Cloudflare (CDN, WAF, etc)
  • Experience with CI/CD or general GitOps deployment pattern experience
  • Has led or significantly contributed to incident response and postmortem processes
  • Working knowledge of AWS core services (networking, IAM, compute) beyond just EKS
Nice to have:
  • Experience evaluating or building AI-assisted ops tooling (MCP, agentic runbooks)
  • OpenTelemetry / observability pipeline design
  • Experience with Terraform

Even if you do not meet all the requirements above, we still encourage you to apply for this position. While we try to be thorough with our prerequisites, not everything about you as a candidate can be condensed into a list of bullet points. What do you have to lose?

Fountain offers an incredibly unique work environment.

We employ a diverse team all over the world. Each Fountaineer is given the freedom to do their best work from wherever they choose. We also understand the importance of in-person connections and hold in-person meetings with your team and meet annually as an organization to build our relationships and focus on the future of moving Fountain Forward.

The benefits we offer in the United States include competitive health plans and a retirement plan. Some Fountain-wide perks offered to all employees across the globe include a flexible vacation policy, paid holidays, monthly lunch stipends, annual allowances for ongoing education related to your profession and career advancement, along with home office, cell phone, and wellness reimbursements. Fountain is a global employer, so some benefit offerings will vary from country to country.

Fountain is proud to be an equal opportunity workplace. We welcome applicants of any educational background, gender identity and expression, sexual orientation, religion, ethnicity, age, socioeconomic status, disability, and veteran status.

By submitting an application, you confirm that you have read our Privacy Policy and agree that we may process and retain your personal data for the purpose of recruitment in accordance with applicable data protection laws.

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Site Reliability Engineer
Senior Site Reliability Engineer

SCALIS • Lisboa

Presencial
EUR 65 000 - 95 000
Senior Data Engineer, Agentic Systems
Senior Data Engineer, Agentic Systems

Fountain • Lisboa

Presencial
EUR 60 000 - 90 000
Flexible vacation
Paid holidays
Lunch stipends
+3
Senior SRE: Platform Reliability & Kubernetes
Senior SRE: Platform Reliability & Kubernetes

SCALIS • Lisboa

Presencial
EUR 65 000 - 95 000
Senior Product Manager - Source - Contract
Senior Product Manager - Source - Contract

Fountain • Lisboa

Teletrabalho
EUR 70 000 - 90 000
Flexible vacation policy
Monthly lunch stipends
Annual education allowance
Senior Site Reliability Engineer: Scale, Resilience & Cloud Ops
Senior Site Reliability Engineer: Scale, Resilience & Cloud Ops

Fountain • Lisboa

Presencial
EUR 70 000 - 110 000
Site Reliability Engineering (Sre)
Site Reliability Engineering (Sre)

Fyld • Lisboa

Presencial
EUR 70 000 - 95 000
Customer Reliability Engineer
Customer Reliability Engineer

Webhosting • Lisboa

Híbrido
EUR 61 000 - 84 000
Equity plan
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Air Apps • Lisboa

Presencial
EUR 70 000 - 110 000
Apple hardware ecosystem
Annual bonus
Health and life insurance
+7
Site Reliability Engineering Manager (Data Infra)
Site Reliability Engineering Manager (Data Infra)

Complyadvantage • Lisboa

Híbrido
EUR 86 000 - 96 000
Equity participation
Private medical insurance
Unlimited Time Off Policy
+2
Global IT Site Reliability Engineer Senior Manager
Global IT Site Reliability Engineer Senior Manager

Boston Consulting Group (BCG) • Lisboa

Presencial
EUR 60 000 - 80 000