Senior SRE: Scale Platform Reliability & Observability

Fountain

Lisboa

Presencial

EUR 70 000 - 110 000

Tempo integral

há 33 horas
Torna-te num dos primeiros candidatos

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Vantagens oferecidas por esta oferta de emprego

Flexible vacation policy
Paid holidays
Education allowance
Home office reimbursements
Wellness reimbursements

Resumo da oferta

Fountain is seeking an experienced Site Reliability Engineer (SRE) to own reliability, scalability, and operational excellence of our platform. You will collaborate with Engineering, Security, and Product to build resilient infrastructure and elevate observability and incident response.

You will design and maintain Kubernetes-based systems, manage AWS workloads, and advance CI/CD and GitOps practices. A strong emphasis on incident management, SLOs, and automation drives this role.

Qualificações

  • 5+ years in site reliability engineering.
  • Own reliability for production systems — define SLOs or run error budgets.
  • Hands-on with Kubernetes/Helm on EKS in production.
  • Experience with CI/CD or GitOps deployment patterns.
  • Led or significantly contributed to incident response and postmortems.

Responsabilidades

  • Own and improve platform reliability (SLOs/SLIs).
  • Design, build, and maintain Kubernetes-based infrastructure and deployment workflows.
  • Operate and evolve our AWS footprint with a security-first mindset.
  • Improve CD/GitOps practices (ArgoCD) and deployment safety.
  • Build autoscaling strategies for services and workloads (KEDA).
  • Lead incident response: on-call, triage, mitigation, postmortems.
  • Strengthen observability across services: metrics, logs, traces, alerts (OpenTelemetry).
  • Partner with application teams to tune performance and reduce toil.
  • Improve IaC practices and manage Terraform modules/environments.
  • Evaluate and integrate AI tooling to accelerate operations.

Conhecimentos

SRE experience
Kubernetes
AWS
CI/CD GitOps
Incident response
Observability

Ferramentas

ArgoCD
Terraform
OpenTelemetry
Cloudflare

Descrição da oferta de emprego

Fountain is seeking an experienced Site Reliability Engineer (SRE) to own reliability, scalability, and operational excellence of our platform. You will collaborate with Engineering, Security, and Product to build resilient infrastructure and elevate observability and incident response.

You will design and maintain Kubernetes-based systems, manage AWS workloads, and advance CI/CD and GitOps practices. A strong emphasis on incident management, SLOs, and automation drives this role.

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Fountain • Lisboa

Presencial
EUR 70 000 - 110 000
Flexible vacation policy
Paid holidays
Education allowance
+2
Senior SRE: Build Resilient Healthcare Data Platform
Senior SRE: Build Resilient Healthcare Data Platform

Promptly Health • Portugal

Presencial
EUR 70 000 - 110 000
Annual performance bonus
Equity compensation via ESOP
Private health insurance
+2
Remote Senior SRE — Scale Global Systems & Automate Resilience
Remote Senior SRE — Scale Global Systems & Automate Resilience

EPAM Systems • Portugal

Presencial
EUR 60 000 - 90 000
Remote Senior AWS SRE: Scale Reliability & Cloud Innovation
Remote Senior AWS SRE: Scale Reliability & Cloud Innovation

ITDS • Portugal

Presencial
EUR 70 000 - 90 000
Remote work flexibility
Career growth opportunities
Remote Platform Security SRE — Kubernetes & IAM
Remote Platform Security SRE — Kubernetes & IAM

Promptly Health • Portugal

Presencial
EUR 90 000 - 130 000
Private health insurance
Equity via ESOP
Annual training allowance
+2
Senior Platform Reliability Engineer: Incident & Automation
Senior Platform Reliability Engineer: Incident & Automation

Arcesium • Lisboa

Presencial
EUR 60 000 - 90 000
SRE Manager: Lead Cloud Reliability & Scale
SRE Manager: Lead Cloud Reliability & Scale

Complyadvantage • Lisboa

Híbrido
EUR 86 000 - 96 000
Equity participation
Private medical insurance
Unlimited Time Off Policy
+2
Site Reliability Engineer — Automation & Resilience
Site Reliability Engineer — Automation & Resilience

Fulcrum Digital Inc • Portugal

Presencial
EUR 50 000 - 80 000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Arcesium • Lisboa

Presencial
EUR 60 000 - 90 000
Senior SRE — Observability & Cloud Reliability (Porto)
Senior SRE — Observability & Cloud Reliability (Porto)

Claranet limited • Porto

Presencial
EUR 45 000 - 65 000