Site Reliability Engineer

FundApps

Braga

Presencial

EUR 42 000 - 62 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

SteelEye, now merged with FundApps, seeks a Site Reliability Engineer to join its engineering team in Portugal. You will support production systems, ensure high availability, and drive reliability improvements across services.

The role suits candidates with production support, cloud operations, or SRE experience, and a strong focus on incident response and automation. Visa sponsorship is not available for this position.

Qualificações

  • Solid Linux systems administration experience and troubleshooting skills.
  • Experience with AWS services and Kubernetes in production environments.
  • Hands-on incident response and runbook familiarity.

Responsabilidades

  • Operate and support the production environment, responding to incidents.
  • Deploy, rollback and contribute to automation of routine tasks.
  • Triage and troubleshoot across services, infra, and network layers.
  • Monitor systems with observability tools and tune SLOs/alerts.
  • Collaborate with platform teams to improve reliability and scalability.

Conhecimentos

Linux administration
AWS (EC2/S3/IAM)
Kubernetes
SRE / Production Support
Incident response
Bash/Python scripting
Communication
Calm under pressure
Learning mindset

Ferramentas

Prometheus
Grafana
ELK
CI/CD pipelines

Descrição da oferta de emprego

SteelEye, now merged with FundApps, provides communications, trade and order surveillance to some of the world’s largest financial institutions.

Our work enhances financial compliance, prevents market abuse, and promotes trust in the financial markets. Our people are passionate about leveraging data and technology to make this happen.

We are looking for a Site Reliability Engineer (SRE) to join our engineering team. You will be responsible for supporting and maintaining our production estate, ensuring availability, performance and reliability across systems and services.

This role is ideal for candidates with prior experience in production support, cloud operations or site reliability roles.

Before you apply please note that that we’re unable to offer visa sponsorship for this role, either now or in the future. Applicants must therefore already have the right to work in Portugal without requiring current or future sponsorship.

Key responsibilities:
  • Operate and support the production environment, responding to incidents and ensuring systems remain highly available.
  • Execute standard operational procedures (e.g. deployments, rollbacks, failovers) and contribute to automation of routine tasks.
  • Triage and troubleshoot production issues across services, infrastructure and network layers.
  • Monitor systems using observability tools, contributing to alert tuning and service level objectives.
  • Collaborate with platform teams to improve reliability, operability, and scalability.
Key skills and experience:
  • Solid understanding of Linux systems administration (troubleshooting, permissions, system services).
  • Basic to intermediate experience with AWS services (e.g., EC2, S3, IAM, EKS).
  • Exposure to Kubernetes (e.g., running pods, reading logs, basic kubectl usage).
  • Hands on experience with production environments, preferable in roles such as SRE, Cloud Support Engineer or Production Support Engineer.
  • Familiarity with incident response and operational run books.
  • Scripting skills in Bash, Python, or similar.
  • Strong communication and collaboration skills.
  • Calm under pressure, particularly during incident response.
  • Eagerness to learn and continuously improve operational excellence.
Nice to Have:
  • Familiarity with CI/CD pipelines and deployment automation.
  • Knowledge of monitoring/logging tools like Prometheus, Grafana and ELK
  • Exposure to security and compliance practices in cloud environments.
Interview Process:

The interview process is structured to assess candidates thoroughly across various competencies and skills relevant to the role. Here's a description of each stage:

  • Intro call with HR
  • First Stage Overview Interview with the team
  • Final Interview with the Lead of Engineering Pillar
Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Production SRE: Reliability, Cloud Ops & Incidents
Production SRE: Reliability, Cloud Ops & Incidents

FundApps • Braga

Presencial
EUR 42 000 - 62 000
Site Reliability Engineer
Site Reliability Engineer

Fulcrum Digital Inc • Lisboa

Presencial
EUR 50 000 - 70 000
Senior Site Reliability Engineer - PSRE
Senior Site Reliability Engineer - PSRE

Arcesium LLC • Lisboa

Híbrido
EUR 50 000 - 80 000
Flexible work arrangements
Competitive compensation and benefits
Continuous learning and development opportunities
Site Reliability Engineer
Site Reliability Engineer

MOZAYDO • Portugal

Híbrido
EUR 30 000 - 40 000
Culture of autonomy and trust
Challenging projects
Opportunities for growth
Site Reliability Engineer
Site Reliability Engineer

Intuition IT – Intuitive Technology Recruitment • Portugal

Presencial
EUR 50 000 - 70 000
Site Reliability Engineer
Site Reliability Engineer

Komodo Consulting • Lisboa

Híbrido
EUR 65 000 - 90 000
Hybrid work model
DevOps Infrastructure Engineer | SRE
DevOps Infrastructure Engineer | SRE

Decskill • Portugal

Presencial
EUR 55 000 - 75 000
Systems Engineer – SRE
Systems Engineer – SRE

Matchtech • Lisboa

Híbrido
EUR 55 000 - 65 000
Private healthcare plan
Meal allowance
Profit-sharing opportunities
+1
Site Reliability Engineer
Site Reliability Engineer

TLScontact • Lisboa

Híbrido
EUR 52 000 - 76 000
Health insurance by Médis
Dental Insurance by Victoria
Flexible work (Hybrid within Portugal)
Site Reliability Engineer
Site Reliability Engineer

Winning • Portugal

Teletrabalho
EUR 45 000 - 65 000