Staff SRE - Global Remote Reliability Leader

Jobgether

España

A distancia

PHP 11.118.000 - 15.075.000

Jornada completa

Hace 5 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Transforma esta oferta en una entrevista: un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Fully remote work environment
First dedicated SRE role
Autonomy to set standards
Career growth in reliability

Descripción de la vacante

Jobgether is seeking a Staff Site Reliability Engineer (remote) to establish organization-wide reliability practices across a fully remote, globally distributed engineering team. You will lead SRE transformation, define SLIs/SLOs, manage incidents, and coach engineers to build scalable, observable systems.

You will drive reliability in multi-region architectures, work with AI-enabled tooling, and shape the culture of failure testing and resilient design across production systems.

Formación

  • 10+ years engineering experience including 3+ years in SRE or reliability roles.
  • Experience owning reliability at platform/organization level.
  • Design and adoption of SLIs/SLOs and error budgets across teams.
  • Strong incident leadership and postmortem-driven improvements.
  • Deep understanding of distributed systems and failure modes.
  • Hands-on with Kubernetes, AWS and modern observability platforms.

Responsabilidades

  • Define and implement SLIs/SLOs for critical production paths.
  • Champion error budgets balancing reliability and delivery.
  • Lead incident management lifecycle including postmortems.
  • Improve alerts, anomaly detection, and runbooks with infra teams.
  • Lead reliability assessments for high-risk changes.
  • Embed SRE principles into engineering culture and AI usage for incidents.
  • Coach engineers to become reliability advocates within teams.
  • Develop repeatable operational standards for readiness and on-call.

Conocimientos

SRE leadership
Incident management
SLIs/SLOs
Error budgets
Go/TypeScript
Infrastructure as code
Kubernetes
AWS
Observability tooling
AI tooling for incidents
Communication
Mentoring/coaching

Herramientas

Kubernetes
AWS
Datadog
Elasticsearch
Redis
DynamoDB
Kafka

Descripción del empleo

Jobgether is seeking a Staff Site Reliability Engineer (remote) to establish organization-wide reliability practices across a fully remote, globally distributed engineering team. You will lead SRE transformation, define SLIs/SLOs, manage incidents, and coach engineers to build scalable, observable systems.

You will drive reliability in multi-region architectures, work with AI-enabled tooling, and shape the culture of failure testing and resilient design across production systems.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Staff Site Reliability Engineer
Staff Site Reliability Engineer

Jobgether • España

A distancia
PHP 11.118.000 - 15.075.000
Fully remote work environment
First dedicated SRE role
Autonomy to set standards
+1
Remote Senior SRE: Reliability & Automation Leader
Remote Senior SRE: Reliability & Automation Leader

QAD • Barcelona

Presencial
EUR 90.000 - 150.000
Senior SRE - Multi-Cloud Reliability (Remote)
Senior SRE - Multi-Cloud Reliability (Remote)

Palo Alto Networks • Madrid

Presencial
EUR 75.000 - 110.000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Stellar Cyber • Banyoles

Presencial
EUR 90.000 - 130.000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+6
Remote Staff SRE - AI-Driven Reliability Architect
Remote Staff SRE - AI-Driven Reliability Architect

Yuno • España

Presencial
EUR 90.000 - 130.000
Remote Work
Home Office Bonus
Stock Options
+2
Remote Senior Staff SRE - Global Cloud Reliability
Remote Senior Staff SRE - Global Cloud Reliability

Palo Alto Networks, Inc. • Madrid

Presencial
EUR 70.000 - 100.000
Senior SRE — Hybrid, Barcelona IoT Platform Reliability
Senior SRE — Hybrid, Barcelona IoT Platform Reliability

Tamarind Intelligence • Barcelona

Híbrido
EUR 55.000 - 70.000
Hybrid work model
Salary 55k-70k€ annually
Comprehensive health insurance
Remote Corporate SRE - Internal Platform Reliability
Remote Corporate SRE - Internal Platform Reliability

Last Minute Group • Madrid

A distancia
EUR 45.000 - 65.000
Shorter working week (36 hours)
Flexible working hours
Training and learning opportunities
+2
Remote SRE Lead: Google Cloud for European Growth
Remote SRE Lead: Google Cloud for European Growth

Shakers • Barcelona

Presencial
EUR 70.000 - 110.000
Site Reliability Engineer - Cloud Reliability (Remote/Flex)
Site Reliability Engineer - Cloud Reliability (Remote/Flex)

AgileEngine • Madrid

Híbrido
EUR 45.000 - 60.000