Site Reliability Engineer

Komodo Consulting

Lisboa

Híbrido

EUR 60 000 - 80 000

Tempo integral

há 9 horas
Torna-te num dos primeiros candidatos

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Komodo Consulting in Lisbon is seeking a Site Reliability Engineer to ensure reliability of Azure-based platforms and drive automation across cloud environments.

You will define SLAs/SLOs, manage incidents, and improve observability with Azure Monitor and Application Insights. Collaboration with data, architecture, and development teams will be key to supporting cloud-native workloads in a hybrid Lisbon office.

Qualificações

  • Extensive experience as SRE or similar reliability role.
  • Strong Azure cloud expertise and automation background.
  • Ability to work cross-functionally with architecture, data and development teams.

Responsabilidades

  • Ensure reliability, availability, and health of cloud and data platforms in Azure.
  • Define and implement SLAs, SLOs, operational models, and DR strategies.
  • Manage incidents, alerts, and improve operational processes.
  • Implement observability with Azure Monitor, Application Insights, and telemetry.
  • Monitor performance, availability, and cost efficiency of platforms.
  • Design and implement automation aligned with existing architecture.
  • Collaborate with data, architecture, and development teams to support cloud-native workloads.
  • Provide operational support for data and AI platforms (Microsoft Fabric, Azure ML).

Conhecimentos

SRE
Azure
Automation
Observability
Telemetry
Data platforms
Cross-functional collaboration
Technical communication

Descrição da oferta de emprego

About Us

Komodo Consulting is a technology and strategy firm specializing in Digital Transformation. Operating in Portugal and Poland, we provide IT Consulting & Nearshore services. We support both public and private sector organizations through two main areas:

  • Consulting — with a focus on strategy, investment analysis, and digital process improvement;
  • IT Team Augmentation — helping clients scale and strengthen their tech teams.
The project

We are seeking a Site Reliability Engineer to work on a project for a Technology Company.

You will have the following responsibilities:
  • Ensure the reliability, availability, and operational health of cloud and data platforms within Azure environments;
  • Define and implement SLAs, SLOs, operational models, and disaster recovery strategies;
  • Manage incidents, alerting systems, and continuously improve operational processes;
  • Implement and maintain observability solutions using Azure Monitor, Application Insights, and telemetry-driven insights;
  • Monitor platform performance, availability, and operational consumption, including cost efficiency;
  • Design and implement automation aligned with existing architecture, promoting engineering over manual operations;
  • Collaborate closely with data, architecture, and development teams to support cloud-native platforms and workloads;
  • Provide operational support for data and AI platforms, including Microsoft Fabric and Azure Machine Learning.
You need to have the following skills/experience:
  • Strong experience as an SRE or in a similar reliability engineering role with a software engineering mindset
  • Solid background in cloud environments, particularly Microsoft Azure;
  • Proven experience in automation, observability, and cloud platform operations;
  • Experience implementing monitoring, alerting, and disaster recovery strategies from scratch;
  • Familiarity with Azure Monitor, Application Insights, and telemetry-based observability practices;
  • Experience working with data platforms and/or AI/ML environments (e.g., Azure Machine Learning, Microsoft Fabric);
  • Ability to work cross-functionally with architecture, data, and development teams;
  • Hands-on, autonomous profile with the ability to structure and improve operational models;
  • Strong technical communication skills.
Location

Hybrid (2 days per week at the office in Lisbon)

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Hybrid SRE — Azure Cloud, Observability & AI Ops
Hybrid SRE — Azure Cloud, Observability & AI Ops

Komodo Consulting • Lisboa

Híbrido
EUR 60 000 - 80 000
Site Reliability Engineer Azure
Site Reliability Engineer Azure

ITDS • Portugal

Híbrido
EUR 40 000 - 60 000
Site Reliability Engineer Azure
Site Reliability Engineer Azure

agap2IT Portugal • Lisboa

Presencial
EUR 55 000 - 75 000
Health insurance
Life and personal accident insurance
Free training and certifications
+4
Senior Site Reliability Engineer - Remote (Portugal, Azure)
Senior Site Reliability Engineer - Remote (Portugal, Azure)

KCS iT • Lisboa

Teletrabalho
EUR 70 000 - 100 000
Site Reliability Engineer (SRE) – Azure
Site Reliability Engineer (SRE) – Azure

Mootiva • Lisboa

Híbrido
EUR 47 000 - 68 000
Site Reliability Engineer
Site Reliability Engineer

Fulcrum Digital Inc • Lisboa

Presencial
EUR 50 000 - 70 000
Site Reliability Engineer
Site Reliability Engineer

La Redoute • Leiria

Presencial
EUR 55 000 - 75 000
SRE Azure
SRE Azure

Adentis Portugal • Lisboa

Presencial
EUR 35 000 - 50 000
Benefícios de saúde para colaboradores e familiares
Equilíbrio entre a vida profissional e pessoal
Formação contínua através de um centro profissional
Remote Site Reliability Engineer - Cloud, CI/CD & Automation
Remote Site Reliability Engineer - Cloud, CI/CD & Automation

Conclusion Lifecycle • Portugal

Presencial
EUR 45 000 - 60 000
Possibility of working remotely
Access to continuous training and certifications
Internal mobility program
Site Reliability Engineer
Site Reliability Engineer

MOZAYDO • Portugal

Híbrido
EUR 30 000 - 40 000
Culture of autonomy and trust
Challenging projects
Opportunities for growth