Site Reliability Engineer

Aubay Portugal

Porto

Híbrido

EUR 45 000 - 75 000

Tempo integral

há 21 horas
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Uma candidatura feita para esta oferta — um currículo e uma carta de apresentação personalizados que vão ao encontro do anúncio.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Health insurance
Training Academy
Career progression
Mental health platform
Team events
Family support
Language platform

Resumo da oferta

Aubay Portugal is seeking a Site Reliability/DevOps Engineer to own reliability and performance of production services in a hybrid Porto setup. You will define SLOs, lead incident response, and advance observability across services, collaborating with development on design, capacity, and automation to prevent toil.

Ideal candidates bring 3+ years in production systems, solid Linux and Kubernetes skills, strong English, and a pragmatic approach to security and tooling.

Qualificações

  • 3+ years of professional experience as Site Reliability Engineer, DevOps Engineer, or backend engineer.
  • Hands-on Kubernetes cluster management, deployment, and troubleshooting.
  • Strong observability/monitoring experience with Grafana, Prometheus, Loki.
  • Solid networking concepts (TCP/IP, HTTP, DNS) and distributed systems knowledge.
  • Scripting/automation in Python or Go with a pragmatic, toil-reducing mindset.
  • Excellent English communication with international teams.

Responsabilidades

  • Own reliability, performance, and operability of production services for high availability and graceful degradation.
  • Define, implement, and steward SLOs and error budgets; drive reliability improvements.
  • Lead incident response for critical production issues and conduct postmortems.
  • Enhance observability with metrics, logs, tracing, dashboards, and alerts.
  • Collaborate with development on system design, capacity planning, and automation to prevent toil.
  • Contribute to Kubernetes-based runtime, focusing on deployment, networking, security baselines, and tooling.

Conhecimentos

Linux fundamentals
Kubernetes
Observability & monitoring
Networking fundamentals
Scripting & automation
English communication

Ferramentas

Grafana
Prometheus
Loki

Descrição da oferta de emprego

At Aubay Portugal, we work daily with some of the largest organisations across sectors such as banking, insurance, telecommunications and energy, helping to build and evolve critical systems that support millions of users. We have been in Portugal since 2007 and are part of an international group with a presence in several European countries. More than technology, we build close relationships and deliver real-impact projects.

What will you do?
  • Own the reliability, performance, and operability of production services on the platform, ensuring high availability and graceful degradation.
  • Define, implement, and steward SLOs and error budgets, driving engineering efforts to continuously improve system reliability.
  • Lead incident response for critical production issues, conduct postmortems, and translate learnings into lasting system improvements.
  • Enhance observability across various services by implementing robust metrics, logging, tracing, dashboards, and actionable alerting systems (e.g., Grafana, Prometheus, Loki).
  • Collaborate closely with development teams on system design, capacity planning, and automation to prevent toil and ensure production readiness.
  • Contribute to the Kubernetes-based runtime environment, focusing on deployment patterns, networking, security baselines, and operational tooling.

At Aubay, our focus is on professionals who bring both technical expertise and the right mindset:

What are we looking for?
  • At least 3 years of professional experience as a Site Reliability Engineer, DevOps Engineer, or in backend engineering, preferably operating production systems at scale.
  • Strong proficiency in Linux fundamentals and hands‑on experience with Kubernetes, including cluster management, deployment, and troubleshooting.
  • Proven experience with observability and monitoring tools such as Grafana, Prometheus, and Loki, including designing and implementing alerting systems.
  • Solid understanding of networking concepts (e.g., TCP/IP, HTTP, DNS) and distributed systems fundamentals.
  • Proficiency in scripting and automation (e.g., Python, Go) with a pragmatic operational mindset towards reducing toil and promoting self-service.
  • Excellent communication skills in English, essential for interacting with various international teams.
  • Alignment with our corporate culture and a commitment to integrity and ethics are highly valued;
  • Strong customer focus, teamwork, task prioritisation, and daily commitment are highly appreciated.
What can we offer?

At Aubay, we invest in your development and well-being:

  • Continuous growth through our Training Academy, with technical and soft skills training, mentoring, tech communities, workshops and in‑person meetups;
  • Close support with regular feedback and a structured career advancement model;
  • Comprehensive health insurance, including dental coverage, as well as access to a mental health platform and well‑being initiatives;
  • Team‑oriented environment, including events and initiatives to promote connection and team spirit;
  • Personal and family life support, from the moment your children are born and throughout their educational path;
  • Access to a language platform with conversational classes in multiple languages;
  • Project‑tailored work model: hybrid (2-3 days a week) in Porto.
What will you find here?
  • A close‑knit and informal environment, where direct communication is preferred and people are valued;
  • Recognition of your contribution, with initiatives that celebrate meaningful moments;
  • Freedom to be yourself, without unnecessary formalities;
  • Teams where knowledge sharing, collaboration and joint growth are a reality.
Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

SRE - Site Reliability Engineer - Kubernetes
SRE - Site Reliability Engineer - Kubernetes

Aubay Portugal • Porto

Híbrido
EUR 60 000 - 90 000
Health insurance
Training Academy
Career advancement
+2
Senior Java/Kotlin Developer
Senior Java/Kotlin Developer

Aubay Portugal • Portugal

Presencial
EUR 60 000 - 90 000
Training Academy
Career advancement
Health insurance (Dental)
+3
Senior Cloud Operations Engineer
Senior Cloud Operations Engineer

Aubay Portugal • Lisboa

Híbrido
EUR 55 000 - 85 000
Training Academy
Health insurance (including dental)
Mental health platform
+4
Application Support Team Lead (French Speaker)
Application Support Team Lead (French Speaker)

Aubay Portugal • Lisboa

Híbrido
EUR 52 000 - 72 000
Health insurance
Dental coverage
Mental health platform
+5
Python Developer
Python Developer

Aubay Portugal • Porto

Híbrido
EUR 55 000 - 85 000
Health insurance
Training academy
Career advancement program
+2
Integration And Production Engineer (French Speaker)
Integration And Production Engineer (French Speaker)

Aubay Portugal • Lisboa

Híbrido
EUR 42 000 - 62 000
Health insurance
Training Academy
Language platform access
+2
Openshift DevOps Engineer (French Speaker) - Porto
Openshift DevOps Engineer (French Speaker) - Porto

Aubay Portugal • Porto

Híbrido
EUR 60 000 - 90 000
Health insurance
Training academy
Career progression
+4
DevOps Engineer (Porto)
DevOps Engineer (Porto)

Aubay Portugal • Porto

Híbrido
EUR 45 000 - 65 000
Health Insurance
Performance Management Cycle
Training Academy
+2
System Administrator (Linux)
System Administrator (Linux)

Aubay Portugal • Porto

Híbrido
EUR 60 000 - 90 000
Training Academy
Health insurance (including dental)
Mental health platform access
+3
React Developer (French Speaker)
React Developer (French Speaker)

Aubay Portugal • Portugal

Presencial
EUR 45 000 - 75 000
Health insurance
Training academy
Career progression
+3