Site Reliability Engineer (Cloud & AI Platforms)

Komodo Consulting

Lisboa

Híbrido

EUR 50 000 - 75 000

Tempo integral

há 13 horas
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Transforma esta função numa entrevista — um currículo e uma carta de apresentação criados à volta do que este empregador procura.

Ultrapassa os filtros ATS

Resumo da oferta

Komodo Consulting is seeking a Site Reliability Engineer (Cloud & AI Platforms) to ensure reliability and performance of cloud and data platforms in Microsoft Azure. You will define operating models, handle incidents, and improve observability across services.

You will collaborate with data, architecture, and development teams to automate operations and support AI/ML workloads, including Microsoft Fabric and Azure ML, in a hybrid Lisbon-based setup.

Qualificações

  • Strong SRE background with software engineering mindset.
  • Deep familiarity with Microsoft Azure cloud environments.
  • Proven ability to design and implement automation for cloud platforms.
  • Experience building and operating observability and monitoring solutions.
  • Hands-on with data/AI platforms (e.g., Azure ML, Fabric).

Responsabilidades

  • Ensure reliability, availability, and health of cloud and data platforms in Azure.
  • Define SLAs/SLOs and disaster recovery strategies.
  • Manage incidents and optimize operational processes.
  • Implement observability using Azure Monitor and Application Insights.
  • Monitor performance, cost, and capacity; drive efficiency.
  • Collaborate with architecture, data, and development teams.
  • Provide operational support for data and AI workloads.

Conhecimentos

SRE experience
Azure
Automation
Observability
Cloud-native

Ferramentas

Azure Monitor
Application Insights
Microsoft Fabric
Azure Machine Learning

Descrição da oferta de emprego

About Us

Komodo Consulting is a technology and strategy firm specializing in Digital Transformation. Operating in Portugal and Poland, we provide IT Consulting & Nearshore services. We support both public and private sector organizations through two main areas:



  • Consulting — with a focus on strategy, investment analysis, and digital process improvement;

  • IT Team Augmentation — helping clients scale and strengthen their tech teams.



About Us

Komodo Consulting is a technology and strategy firm specializing in Digital Transformation. Operating in Portugal and Poland, we provide IT Consulting & Nearshore services. We support both public and private sector organizations through two main areas:



  • Consulting — with a focus on strategy, investment analysis, and digital process improvement;

  • IT Team Augmentation — helping clients scale and strengthen their tech teams.



The project

We are seeking a Site Reliability Engineer (Cloud & AI Platforms) to work on a project for a Technology Company.



You will have the following responsibilities:


  • Ensure the reliability, availability, and operational health of cloud and data platforms within Azure environments;

  • Define and implement SLAs, SLOs, operational models, and disaster recovery strategies;

  • Manage incidents, alerting systems, and continuously improve operational processes;

  • Implement and maintain observability solutions using Azure Monitor, Application Insights, and telemetry-driven insights;

  • Monitor platform performance, availability, and operational consumption, including cost efficiency;

  • Design and implement automation aligned with existing architecture, promoting engineering over manual operations;

  • Collaborate closely with data, architecture, and development teams to support cloud-native platforms and workloads;

  • Provide operational support for data and AI platforms, including Microsoft Fabric and Azure Machine Learning.



You need to have the following skills/experience:


  • Strong experience as an SRE or in a similar reliability engineering role with a software engineering mindset

  • Solid background in cloud environments, particularly Microsoft Azure;

  • Proven experience in automation, observability, and cloud platform operations;

  • Experience implementing monitoring, alerting, and disaster recovery strategies from scratch;

  • Familiarity with Azure Monitor, Application Insights, and telemetry-based observability practices;

  • Experience working with data platforms and/or AI/ML environments (e.g., Azure Machine Learning, Microsoft Fabric);

  • Ability to work cross-functionally with architecture, data, and development teams;

  • Hands-on, autonomous profile with the ability to structure and improve operational models;

  • Strong technical communication skills.



Location

Hybrid (2 days per week at the office in Lisbon)

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Claranet • Lisboa

Híbrido
EUR 55 000 - 85 000
Regular professional development
Certification paths resources
Regular teambuilding programs
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Claranet Portugal • Portugal

Presencial
EUR 45 000 - 65 000
Cloudops Engineer
Cloudops Engineer

Sgi • Setúbal

Presencial
EUR 55 000 - 85 000
Azure Site Reliability Engineer – Cloud & DevOps (Hybrid, Lisbon)
Azure Site Reliability Engineer – Cloud & DevOps (Hybrid, Lisbon)

act digital • Lisboa

Híbrido
EUR 80 000 - 100 000
Hybrid work regime
Senior Site Reliability Engineer
Senior Site Reliability Engineer

outsystems • Portugal

Híbrido
EUR 60 000 - 90 000
Site Reliability Engineer Azure
Site Reliability Engineer Azure

agap2IT Portugal • Lisboa

Presencial
EUR 60 000 - 90 000
Health insurance
Life and personal accident insurance
Free training and certifications
+3
Senior Cloud Engineer
Senior Cloud Engineer

login.works • Lisboa

Híbrido
EUR 60 000 - 90 000
Collaborative work culture
Exposure to cutting-edge AI topics
Opportunity to work on impactful products
+1
Site Reliability Engineer
Site Reliability Engineer

La Redoute • Viseu

Presencial
EUR 60 000 - 86 000
Senior Cloud Engineer
Senior Cloud Engineer

Legal Tech Organization • Lisboa

Híbrido
EUR 70 000 - 90 000
Hybrid role
Office in Lisbon (20% onsite)
Site Reliability Engineer (SRE) @Lisboa
Site Reliability Engineer (SRE) @Lisboa

KCS iT • Lisboa

Híbrido
EUR 40 000 - 60 000
Programmes de formation gratuits
Expérience internationale
Options de travail flexible (hybride, à distance, sur site)
+1