Site Reliability Engineer

Capgemini

Vila Nova de Gaia

Presencial

EUR 40 000 - 70 000

Tempo integral

Há 2 dias
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Transforma esta função numa entrevista — um currículo e uma carta de apresentação criados à volta do que este empregador procura.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Health and Life insurance
Referral program with bonuses
Career Acceleration Programs

Resumo da oferta

Capgemini Portugal is building a Tech Hub in Porto and seeks a Site Reliability Engineer to improve availability, performance and resilience of cloud-native platforms. You will own observability and incident response, driving RCAs and post-incident reviews.

You will collaborate with Engineering, Product, Security and Data teams to evolve reliability frameworks, automate operations, and support CI/CD deployments in a hybrid work setup.

Qualificações

  • Strong understanding of SRE principles and production operations.
  • Experience supporting high-availability production environments.
  • Proven incident management and root cause analysis skills.
  • Experience with Linux-based environments and system administration.
  • Hands-on containerization with Docker and Kubernetes.
  • Knowledge of observability and monitoring tools (Prometheus, Grafana, ELK, Splunk).
  • Experience building CI/CD pipelines (GitHub Actions, Azure DevOps).
  • Familiarity with IaC (Terraform, Ansible).
  • Solid understanding of networking fundamentals and cloud-native architectures.

Responsabilidades

  • Ensure availability, performance, reliability and resilience of cloud-native platforms.
  • Own and improve platform observability with metrics, logs, traces, dashboards, alerts.
  • Define and monitor SLIs/SLOs/SLAs with Product and Engineering.
  • Drive incident response, root cause analysis, and post-incident reviews.
  • Identify bottlenecks and implement reliability improvements.
  • Collaborate with Engineering, Product, Architecture, Security, Data to evolve reliability.
  • Develop automation to reduce operational overhead and improve resilience.
  • Support CI/CD pipelines, IaC, and deployment processes.
  • Participate in on-call rotations and production support.
  • Design and operate distributed systems on cloud platforms.

Conhecimentos

SRE principles
On-call support
Linux administration
Docker
Kubernetes
Observability
Prometheus
Grafana
ELK Stack
Splunk
GitHub Actions
Azure DevOps
Terraform
Ansible
Networking basics
Cloud-native architectures
Agile environments
Incident/Problem/Change mgmt

Ferramentas

Docker
Kubernetes
GitHub Actions
Azure DevOps
Terraform
Ansible
Prometheus
Grafana
ELK Stack
Splunk

Descrição da oferta de emprego

Site Reliability EngineerChoosing Capgemini means choosing a company where you will be empowered to shape your career in the way you’d like, where you’ll be supported and inspired by a collaborative community of colleagues around the world, and where you’ll be able to reimagine what’s possible. Join us and help the world’s leading organizations unlock the value of technology and build a more sustainable, more inclusive world.You’ll have the opportunity to collaborate with a leading European mobility services provider, headquartered in Germany and backed by nearly 90 years of expertise. As part of this journey, you’ll contribute to the launch of a new Tech Hub in Porto, helping to drive innovative, technology-led solutions that are transforming how businesses move while contributing to a more sustainable future.YOUR ROLEEnsure the availability, performance, reliability, and resilience of cloud-native platforms and customer-facing services.Own and continuously improve platform observability through metrics, logs, traces, dashboards, and alerting.Define, implement, and monitor SLIs, SLOs, and SLAs in partnership with Product and Engineering teams.Drive incident response, root cause analysis, and post-incident reviews to improve platform stability.Identify performance bottlenecks and proactively implement reliability and scalability improvements.Embed reliability engineering practices throughout the Software Development Life Cycle (SDLC).Collaborate with Engineering, Product, Architecture, Security, and Data teams to establish and evolve reliability frameworks.Develop and maintain automation solutions to reduce operational overhead and improve system resilience.Support and optimize CI/CD pipelines, infrastructure automation, and deployment processes.Participate in on-call rotations and support production environments, ensuring rapid resolution of critical issues.Contribute to the design and operation of distributed systems running on modern cloud platforms.YOUR PROFILEStrong understanding of Site Reliability Engineering (SRE) principles, production operations, and reliability best practices.Experience supporting and operating high-availability production environments, including on-call support.Proven incident management, troubleshooting, and root cause analysis skills in complex distributed systems.Experience working with Linux-based environments and system administration.Hands-on experience with containerization and orchestration technologies such as Docker and Kubernetes.Strong knowledge of observability and monitoring tools, including Prometheus, Grafana, ELK Stack, and Splunk.Experience building and maintaining CI/CD pipelines using tools such as GitHub Actions and Azure DevOps.Familiarity with Infrastructure as Code (IaC) and configuration management tools, including Terraform and Ansible.Solid understanding of networking fundamentals, cloud-native architectures, and distributed systems.Experience collaborating with development, infrastructure, and security teams within Agile environments.Knowledge of incident, problem, and change management processes.WHAT YOU'LL LOVE ABOUT WORKING HEREAt Capgemini Portugal we have a flexible and dynamic work environment. Flexibility enables a better work-life balance and gives more flexibility to the employee to manage the working hours, as well if he works at the office or remotely, according with the company’s hybrid work policy;We have local programs that promote people growth, reskill and new skills development (Career Acceleration Programs);We promote an empowering environment with autonomy and peers' relationships among the top scores of our Monthly Employees' feedback;Next to this, we also offer an attractive compensation package and benefits such as Health and Life insurance, as well as Referral program with bonuses for talent recommendations and other fringe benefits according with our partnerships in for;Capgemini Portugal is an equal opportunity employer. We promote equality and dignity in all aspects of recruitment and employment, as well as employment offers and promotions made according with competence and ability or performance, respectively.ABOUT CAPGEMINICapgemini is a global business and technology transformation partner, helping organizations to accelerate their dual transition to a digital and sustainable world, while creating tangible impact for enterprises and society. It is a responsible and diverse group of 340,000 team members in more than 50 countries. With its strong over 55-year heritage, Capgemini is trusted by its clients to unlock the value of technology to address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fuelled by its market leading capabilities in AI, cloud and data, combined with its deep industry expertise and partner ecosystem. The Group reported 2023 global revenues of €22.5 billionGet the future you want | www.capgemini.com
Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

Junior Site Reliability Engineer
Junior Site Reliability Engineer

Capgemini • Vila Nova de Gaia

Presencial
EUR 32 000 - 48 000
Health and Life insurance
Referral program with bonuses
Hybrid work policy
Senior Fullstack Developer
Senior Fullstack Developer

Capgemini • Vila Nova da Barquinha

Presencial
EUR 55 000 - 75 000
Health and life insurance
Referral program with bonuses
Devops Engineer (Porto)
Devops Engineer (Porto)

Capgemini • Vila Nova da Barquinha

Híbrido
EUR 45 000 - 65 000
Health and life insurance
Referral program with bonuses
Senior Fullstack Developer (Typescript & React / SRE)
Senior Fullstack Developer (Typescript & React / SRE)

Capgemini • Vila Nova de Gaia

Presencial
EUR 50 000 - 70 000
Health insurance
Life insurance
Referral bonuses
Senior Expert Information Security
Senior Expert Information Security

Capgemini • Vila Nova de Gaia

Presencial
EUR 55 000 - 75 000
Health and Life insurance
Referral program with bonuses
Career Acceleration Programs
+1
Junior Fullstack Developer (Typescript & React) (Porto)
Junior Fullstack Developer (Typescript & React) (Porto)

Capgemini • Vila Nova de Gaia

Presencial
EUR 40 000 - 60 000
Health insurance
Referral program with bonuses
Solution Architect
Solution Architect

Capgemini • Vila Nova de Gaia

Presencial
EUR 65 000 - 90 000
Health and Life insurance
Hybrid work policy
Career Acceleration Programs
+1
Agile Master (Porto)
Agile Master (Porto)

Capgemini • Vila Nova de Gaia

Presencial
EUR 42 000 - 66 000
Health and Life insurance
Referral program with bonuses
Senior Fullstack Developer
Senior Fullstack Developer

Capgemini • Entroncamento

Híbrido
EUR 50 000 - 75 000
Flexible work model
Career acceleration programs
Autonomy and collaboration
+2
Senior IT Security Officer
Senior IT Security Officer

Capgemini • Lisboa

Híbrido
EUR 70 000 - 95 000
Health insurance
Life insurance
Referral program
+1