Observability Engineer

LUZA Group

Leiria

Remote

EUR 45,000 - 65,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Remote work when possible
Equipment provided
Benefits plan

Job summary

LUZA Group is seeking a Site Reliability Engineer (SRE) with strong observability expertise to scale our new observability stack across platforms. You will own metrics, logs, traces, and RUM pipelines, and collaborate with cross-functional teams to embed observability by design.

The role emphasizes automation, cost awareness, and a product-oriented mindset with a You build it, you run it culture. Fluency in English is required and most work is remote.

Qualifications

  • 3+ years as an SRE/Observability Engineer.
  • Experience with OpenTelemetry.
  • English fluency.

Responsibilities

  • Design, implement, and maintain observability solutions covering metrics, logs, traces, and RUM;
  • Work with Grafana Cloud, Tempo, Loki, Mimir, Alloy, and OpenTelemetry;
  • Build reliable alerting and monitoring pipelines based on SLOs/SLAs;
  • Ensure health and integrity of observability data flows from instrumentation to dashboards;
  • Collaborate with development and operations teams to integrate observability into the software lifecycle;
  • Define and promote observability standards across the organization;
  • Support modernization by evolving legacy monitoring and alerting;
  • Monitor observability costs and contribute to FinOps;
  • Mostly remote.

Skills

SRE experience
Observability mindset
English fluency

Tools

Kubernetes
Helm
Terraform
ArgoCD
OpenTelemetry
Grafana Cloud
Tempo
Loki
Mimir
Alloy
Java instrumentation

Job description

Luza is an all-inclusive consulting company focused on talent, tech and innovation. We exist to elevate companies and humans all around the world, making change, from the inside to the outside.

We believe that technology + human kindness positively impacts every community around the world. Our approach is simple, we see a world without borders, and believe in equal opportunities. We are guided by our core principles of spreading positivity, good energy and promote equality and care for others.

Our hiring process is unique! People are selected by their value, education, talent and personality. We dont present ethnicity, religion, national origin, age, gender, sexual orientation or identity.

Its time to burst the bubble, and we will do it together!

What You'll do:

We are looking for a Site Reliability Engineer (SRE) with strong expertise in Observability Engineering to join our team. This role is pivotal to ensuring the reliability, visibility, and performance of our platforms and services. The ideal candidate will have hands-on experience with the Grafana Stack (Tempo, Loki, Mimir, Alloy), knowledge in Java development, a strong SRE mindset, and a passion for automation, scalability, and ownership. You'll be joining a motivated, cross-functional team responsible for implementing and scaling our new observability stack that is being built to be a platform for observability for the whole company. Your contributions will directly impact system performance, user experience, and operational cost efficiency.

  • Design, implement, and maintain observability solutions covering metrics, logs, traces, and RUM;
  • Work with tools such as Grafana Cloud, Tempo, Loki, Mimir, Alloy, and OpenTelemetry;
  • Build reliable alerting and monitoring pipelines based on SLOs/SLAs, focusing on low-maintenance automation;
  • Ensure the health and integrity of observability data flows from instrumentation to dashboards;
  • Collaborate with development and operations teams to embed observability by design into the software lifecycle;
  • Define and promote best practices and standards for observability across the organization;
  • Support the modernization of observability by replacing and evolving legacy monitoring and alerting solutions;
  • Monitor observability-related costs and contribute to FinOps efforts by identifying optimization opportunities;
  • Mosty Remote;
Who You Are:
  • 3+ years of experience as an SRE, Observability Engineer, or equivalent role;
  • Practical experience with OpenTelemetry, or similar instrumentation tools;
  • Experience in Kubernetes, Helm, Terraform, and ArgoCD;
  • Experience designing and managing telemetry pipelines (metrics/logs/traces), exporters, and sidecars;
  • Product-oriented mindset with a bias for automation and a you build it, you run it culture;
  • Fluency in English.
Nice-to-have:
  • Knowledge of APM and distributed tracing solutions;
  • Experience with FinOps practices applied to observability;
  • Hands-on involvement in replacing legacy monitoring stacks;
  • Experience with Cloud environments (Azure preferred);
  • Contributions to open-source observability tools;
  • Knowledge in Java development and applications instrumentation;
  • Expertise in performance monitoring, alerting, dashboarding, and root cause analysis.
What you'll get:
  • Wage according to candidate's professional experience;
  • Remote Work whenever possible;
  • Delivery of work equipment adjusted to the performance of functions;
  • Benefits plan;
  • And others.

Work together with expert teams on projects of large magnitude and intensity, long term together with our clients, all leaders in their industries.

Are you ready to step into a diverse and inclusive world with us?

Together we will promote uniquess!
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer – Observability
Site Reliability Engineer – Observability

Decskill • Portugal

Remote
EUR 55,000 - 75,000
Long-term projects
Growth opportunities
People-first culture
+1
Observability Engineer
Observability Engineer

Luza Group • Lisboa

On-site
EUR 40,000 - 70,000
Remote work flexibility
Equipment provided
Benefits plan
Observability & SRE Engineer - Remote
Observability & SRE Engineer - Remote

LUZA Group • Leiria

Remote
EUR 45,000 - 65,000
Remote work when possible
Equipment provided
Benefits plan
Remote Site Reliability Engineer – Observability
Remote Site Reliability Engineer – Observability

Decskill • Portugal

Remote
EUR 55,000 - 75,000
Long-term projects
Growth opportunities
People-first culture
+1
Senior Observability Engineer
Senior Observability Engineer

GRiT Solutions • Lisboa

On-site
EUR 65,000 - 90,000
Continuous training
GRiT prizes
Partnership discounts
+4
Software Engineer I - AI Observability
Software Engineer I - AI Observability

Elastic • Portugal

Remote
EUR 39,000 - 61,000
Health coverage for you and family
Flexible locations and schedules
Generous vacation days
+1
Systems Engineer – SRE
Systems Engineer – SRE

Matchtech • Lisboa

On-site
EUR 55,000 - 65,000
Private healthcare plan
Meal allowance
Profit-sharing opportunities
+1
Senior Fullstack Java & React Developer
Senior Fullstack Java & React Developer

LUZA Group • Lisboa

Remote
EUR 50,000 - 75,000
Remote work
Health insurance
Work equipment provided
Senior Elasticsearch & Observability Engineer
Senior Elasticsearch & Observability Engineer

bridge351 • Lisboa

Hybrid
EUR 70,000 - 90,000
Health insurance
Life insurance
Tech visa support
Software Engineer I - AI Observability
Software Engineer I - AI Observability

AI Chopping Block, Inc. • Portugal

Remote
EUR 39,000 - 61,000
Health coverage for you and family
Flexible locations and schedules
Generous vacation days
+3