Senior Site Reliability Engineer

Claranet limited

Lisboa

On-site

EUR 60,000 - 90,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Claranet limited is seeking a Site Reliability Engineer to join our team in Lisbon, Portugal. The role blends operational excellence with engineering projects to build reliable, scalable, and secure cloud platforms.

You will work on Azure-based infrastructure and Kubernetes environments, focusing on automation, observability, and continuous improvement of service reliability. Strong scripting and English communication are essential.

Qualifications

  • Degree in Computer Science, Engineering or a related field, or equivalent practical experience.
  • At least 4 years of experience in SRE, DevOps, Platform Engineering, Cloud Engineering or a similar role.
  • Hands-on experience with Microsoft Azure, particularly compute, networking and storage services.
  • Practical experience with Kubernetes; experience with AKS is an advantage.
  • Experience with Terraform or another Infrastructure as Code tool.
  • Familiarity with CI/CD practices and version control systems.
  • Experience with monitoring, logging and alerting platforms such as Datadog, Azure Monitor, Prometheus, Grafana or equivalent.
  • Good scripting skills in Bash, Python or PowerShell.
  • Understanding of software development and deployment practices.
  • Experience with .NET and/or Java, microservices or business applications deployed on Kubernetes is a strong advantage.
  • Ability to troubleshoot complex technical issues in a structured and collaborative way.
  • Good written and verbal communication skills in English.

Responsibilities

  • Monitoring and maintaining cloud and Kubernetes platforms to ensure high availability and performance.
  • Investigating and resolving incidents, conducting root cause analysis, and driving continuous service improvements.
  • Designing, deploying, and managing scalable infrastructure in Azure.
  • Managing Kubernetes environments, preferably with AKS (Azure Kubernetes Service).
  • Developing and maintaining Infrastructure as Code using Terraform.
  • Building and improving CI/CD pipelines and automating operational processes.
  • Implementing observability solutions, including monitoring, logging, tracing, and alerting tools such as Datadog.
  • Defining and tracking reliability and performance metrics (SLIs, SLOs, and error budgets).
  • Collaborating with development and infrastructure teams to deliver reliable, secure, and maintainable platform solutions.
  • Promoting DevOps, automation, knowledge sharing, and a culture of continuous improvement.

Skills

SRE
DevOps
Platform Engineering
Cloud Engineering
Azure
Kubernetes
AKS
Terraform
CI/CD
Datadog
Azure Monitor
Prometheus
Grafana
Bash
Python
PowerShell
.NET
Java

Education

Degree in Computer Science or Engineering

Tools

Terraform
AKS
Datadog
Azure Monitor
Prometheus
Grafana

Job description

We're fast learners,hard workers, natural collaborators... and we Make Modern Happen!

Our ambition is tounlock the potential of our digital world so that organisations everywhere caninnovate and thrive securely.

We aim to achievethis goal by bringing together the world’s most talented people and the mostpowerful technologies, combining them to address our customers' challenges and tobuild something stronger together.

We are looking for a SiteReliability Engineer to join our team and help us build and operatereliable, scalable and secure technology platforms.

This role combinestwo complementary areas of work:

  • 50% Operational Excellence: ensuring the smoothoperation of our platforms, responding to service requests and incidents,troubleshooting issues and continuously improving reliability.
  • 50% Engineering & ImprovementProjects: designing and implementing automation, observability, infrastructure andplatform improvements that make our services more resilient and easier tooperate.

This role isresponsible for ensuring the reliability, performance, security, andscalability of cloud-based platforms, primarily in Azure and Kubernetes environments. The position combines operational support, infrastructureengineering, automation, and Site Reliability Engineering (SRE) practices.

Your responsibilities include:
  • Monitoring andmaintaining cloud and Kubernetes platforms to ensure high availability andperformance.
  • Investigating andresolving incidents, conducting root cause analysis, and drivingcontinuous service improvements.
  • Designing, deploying,and managing scalable infrastructure in Azure.
  • Managing Kubernetesenvironments, preferably with AKS (Azure Kubernetes Service).
  • Developing andmaintaining Infrastructure as Code using Terraform.
  • Building and improvingCI/CD pipelines and automating operational processes.
  • Implementingobservability solutions, including monitoring, logging, tracing, andalerting tools such as Datadog.
  • Defining and trackingreliability and performance metrics (SLIs, SLOs, and error budgets).
  • Collaborating withdevelopment and infrastructure teams to deliver reliable, secure, andmaintainable platform solutions.
  • Promoting DevOps,automation, knowledge sharing, and a culture of continuous improvement.
You must have:
  • Degree in Computer Science, Engineering or a related field, orequivalent practical experience.
  • At least 4 years of experience in SRE, DevOps, PlatformEngineering, Cloud Engineering or a similar role.
  • Hands-on experience with Microsoft Azure, particularly compute,networking and storage services.
  • Practical experience with Kubernetes; experience with AKS is anadvantage.
  • Experience with Terraform or another Infrastructure as Code tool.
  • Familiarity with CI/CD practices and version control systems.
  • Experience with monitoring, logging and alerting platforms such asDatadog, Azure Monitor, Prometheus, Grafana or equivalent.
  • Good scripting skills in Bash, Python or PowerShell.
  • Understanding of software development and deployment practices.
  • Experience with .NET and/or Java, microservices or businessapplications deployed on Kubernetes is a strong advantage.
  • Ability to troubleshoot complex technical issues in a structuredand collaborative way.
  • Good written and verbal communication skills in English.
We value:
  • Experience with AWS or Google Cloud.
  • Experience with .NET or Java application development.
  • Knowledge of SLI/SLO frameworks, error budgets and incidentmanagement practices.
  • Experience with distributed systems, APIs and cloud-nativearchitectures.
  • Familiarity with security, networking and identity concepts inAzure and Kubernetes.
  • Relevant certifications, such as:
  • MicrosoftCertified: Azure Solutions Architect Expert
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer /Hybrid
Site Reliability Engineer /Hybrid

A2IT Technology • Leiria

On-site
EUR 52,000 - 76,000
05 - Site Reliability Engineer
05 - Site Reliability Engineer

PPM Coachers • Lisboa

On-site
EUR 60,000 - 90,000
05 - Site Reliability Engineer
05 - Site Reliability Engineer

PPM Coachers • Lisboa

On-site
EUR 55,000 - 85,000
Site Reliability Engineer (SRE) M/F
Site Reliability Engineer (SRE) M/F

Ankix • Portugal

On-site
EUR 50,000 - 75,000
Site Reliability Engineering Manager (Data Infra)
Site Reliability Engineering Manager (Data Infra)

Complyadvantage • Lisboa

On-site
EUR 86,000 - 96,000
Equity participation
Private medical insurance
Unlimited Time Off Policy
+2
Site Reliability Engineer
Site Reliability Engineer

TEKEVER • Lisboa

On-site
EUR 60,000 - 90,000
Excellent work environment
Flexible work arrangements
Professional development opportunities
+1
Site Reliability Engineer (SRE) – Azure
Site Reliability Engineer (SRE) – Azure

Mootiva • Lisboa

On-site
EUR 47,000 - 68,000
Senior Site Reliability Engineer - Platform Reliability (Resilience)
Senior Site Reliability Engineer - Platform Reliability (Resilience)

Elastic • Portugal

Remote
EUR 63,000 - 84,000
Site Reliability Engineer (SRE) – Azure
Site Reliability Engineer (SRE) – Azure

Mootiva • Leiria

On-site
EUR 50,000 - 80,000
Senior Platform Engineer – Azure
Senior Platform Engineer – Azure

Decskill • Portugal

Remote
EUR 70,000 - 110,000
Remote working model
International career opportunities