Lead SRE - Global Reliability (Remote)

Jobgether

Portugal

Teletrabalho

EUR 120 000 - 180 000

Tempo integral

Há 5 dias
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Recebe uma resposta deste empregador — um currículo e uma carta de apresentação adaptados exatamente ao que estão a contratar.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Fully remote
First dedicated SRE
Autonomy to shape practices
Global collaboration

Resumo da oferta

Jobgether is seeking a Staff Site Reliability Engineer based in Portugal to lead reliability across a global remote engineering organization. You will be the first dedicated SRE, embedding reliability practices and shaping AI-assisted incident analysis.

Responsibilities include defining SLIs/SLOs, leading incident management, and coaching engineers to adopt robust, scalable reliability standards across teams while maintaining strong collaboration with leadership and product teams.

Qualificações

  • 10+ years of engineering experience with at least 3 in SRE or reliability roles across multiple teams.
  • Own reliability at platform or organizational level, not just individual services.
  • Design/implement SLIs, SLOs, and error budgets with cross‑team adoption.
  • Lead high-severity incidents and postmortems that drive measurable improvements.
  • Deep understanding of distributed systems and failure modes.
  • Hands-on with Kubernetes, AWS and modern observability tools (Datadog).
  • Read/write production code in Go or TypeScript; IaC proficiency.
  • Ability to influence teams and establish reliability practices across an org.

Responsabilidades

  • Define and implement SLIs/SLOs for critical production paths and tie to engineering decisions.
  • Champion error budgets balancing reliability with feature delivery across teams.
  • Maintain reliability metrics used by leadership to guide investments.
  • Strengthen incident management lifecycle including postmortems and follow-ups.
  • Improve alerts, anomaly detection, and operational tooling with infra teams.
  • Lead reliability assessments for changes and new services with readiness plans.
  • Embed SRE principles into engineering culture and coach engineers.
  • Develop repeatable operational standards for runbooks and on-call practices.
  • Work with architects to bake reliability into system design and architecture.

Conhecimentos

Kubernetes
AWS
Datadog
Go
TypeScript
Infrastructure as code
Incident leadership
Coaching
Communication
AI tools for incident investigation
Asynchronous decision-making
Distributed systems
Elasticsearch
Redis
DynamoDB
Kafka
FinOps

Ferramentas

Kubernetes
AWS
Datadog
Elasticsearch
Redis
DynamoDB
Kafka

Descrição da oferta de emprego

Jobgether is seeking a Staff Site Reliability Engineer based in Portugal to lead reliability across a global remote engineering organization. You will be the first dedicated SRE, embedding reliability practices and shaping AI-assisted incident analysis.

Responsibilities include defining SLIs/SLOs, leading incident management, and coaching engineers to adopt robust, scalable reliability standards across teams while maintaining strong collaboration with leadership and product teams.

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior SRE — Remote, Cloud-Native & Observability
Senior SRE — Remote, Cloud-Native & Observability

Jobgether • Portugal

Teletrabalho
EUR 110 000 - 150 000
Remote work
Equity
Flexible PTO
+3
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Jobgether • Portugal

Teletrabalho
EUR 120 000 - 180 000
Fully remote
First dedicated SRE
Autonomy to shape practices
+1
Lead SRE: Drive Agile Reliability & DevOps Excellence
Lead SRE: Drive Agile Reliability & DevOps Excellence

Decskill • Portugal

Presencial
EUR 50 000 - 70 000
Remote Site Reliability Engineer - Cloud, CI/CD & Automation
Remote Site Reliability Engineer - Cloud, CI/CD & Automation

Conclusion Lifecycle • Portugal

Presencial
EUR 45 000 - 60 000
Possibility of working remotely
Access to continuous training and certifications
Internal mobility program
Senior Site Reliability Engineer — Hybrid Porto
Senior Site Reliability Engineer — Hybrid Porto

Aubay Portugal • Porto

Híbrido
EUR 60 000 - 90 000
Health insurance
Training Academy
Career advancement
+2
Senior Site Reliability Engineer (Hybrid, Portugal)
Senior Site Reliability Engineer (Hybrid, Portugal)

OutSystems • Lisboa

Híbrido
EUR 70 000 - 110 000
Remote Site Reliability Engineer: Resilience & Observability
Remote Site Reliability Engineer: Resilience & Observability

Intermedia Intelligent Communications • Portugal

Presencial
EUR 40 000 - 70 000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

outsystems • Portugal

Híbrido
EUR 60 000 - 90 000
Production SRE: Reliability, Cloud Ops & Incidents
Production SRE: Reliability, Cloud Ops & Incidents

FundApps • Braga

Presencial
EUR 42 000 - 62 000
SRE: Build Resilient, Scalable Systems (Lisbon, Hybrid)
SRE: Build Resilient, Scalable Systems (Lisbon, Hybrid)

MOZAYDO • Portugal

Híbrido
EUR 30 000 - 40 000