System Reliability Engineer (Application Support + Automation)

Fulcrum Digital

Lisboa

Presencial

EUR 55 000 - 75 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Fulcrum Digital is seeking a System Reliability Engineer to oversee production environments and drive automation initiatives. You will implement monitoring, incident response, and CI/CD automation across multiple geographies as part of a global engineering team.

The role emphasizes automating everything, optimizing system reliability, and collaborating across the tech stack to ensure platform health and faster recovery times.

Qualificações

  • Proficiency in shell scripting and ITIL/ITSM practices.
  • Experience with SQL and application troubleshooting.
  • Familiarity with monitoring tools (Splunk or Dynatrace preferred).
  • CI/CD with Jenkins and automation using Ansible.
  • Groovy scripting and YAML; version control with Git.

Responsabilidades

  • Plan, manage, and oversee all aspects of a production environment.
  • Define strategies for application performance monitoring and optimization.
  • Respond to incidents and incrementally reduce incident counts.
  • Support deployment of code across lower environments and automate processes.
  • Design, develop, and standardize monitoring and alerting mechanisms.

Conhecimentos

Shell scripting
ITIL/ITSM
SQL
Application Troubleshooting
Monitoring tools
Jenkins
Ansible
Groovy/Yaml
Git
Payments Flows

Descrição da oferta de emprego

System Reliability Engineer (Application Support + Automation)

Who are we Fulcrum Digital is an agile and next-generation digital accelerating companyproviding digital transformation and technology services right from ideation toimplementation. These services have applicability across a variety ofindustries, including banking & financial services, insurance, retail,higher education, food, healthcare, and manufacturing.

The Role
  • Plan, manage,and oversee all aspects of a Production Environment
  • Definestrategies for Application Performance Monitoring, Optimization in Prodenvironment
  • Respond toIncidents and improvise platform based on feedback and measure thereduction of incidents over time.
  • Supportdeployment of code into multiple lower environments. Supportingcurrent processes with an emphasis on automating everything as soon aspossible.
  • Design, develop and standardize Monitoring andAlerting mechanism for the supported applications.
  • Take aholistic approach to problem solving, by connecting the dots during aproduction event through the various technology stack that makes up theplatform, to optimize meantime to recover.
  • Engage in andimprove the whole lifecycle of services—from inception and design, throughdeployment, operation and refinement.
  • Analyse ITSMactivities of the platform and provide feedback loop to development teamson operational gaps or resiliency concerns.
  • Supportservices before they go live through activities such as system designconsulting, capacity planning and launch reviews.
  • Support theapplication CI/CD pipeline for promoting software into higher environmentsthrough validation and operational gating, and lead in DevOps automationand best practices.
  • Maintainservices once they are live by measuring and monitoring availability,latency, and overall system health.
  • Scale systemssustainably through mechanisms like automation and evolving systems bypushing for changes that improve reliability and velocity.
  • Work with aglobal team spread across tech hubs in multiple geographies and timezones.
  • Ability toshare knowledge and explain processes and procedures to others.
  • Able toperform on-call duties on a rotational basis.
  • Occasionaloff hours work required.
Requirements
  • ShellScripting
  • ITIL / ITSM
  • SQL - good to have
  • ApplicationTroubleshooting
  • AnyMonitoring tool (Preferred Splunk/Dynatrace)
  • Jenkins -CI/CD
  • Ansible
  • GroovyScripting/Yaml
  • Git basic/bitbucket
Good To Have
  • Payments Flows, Switching, Settlements, Authorisation flows.
Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

System Reliability Engineer (Application Support + Automation)
System Reliability Engineer (Application Support + Automation)

Fulcrum Digital Inc • Lisboa

Presencial
EUR 42 000 - 65 000
Site Reliability Engineer
Site Reliability Engineer

Fulcrum Digital Inc • Portugal

Presencial
EUR 50 000 - 80 000
System Reliability Engineer — App Support & Automation
System Reliability Engineer — App Support & Automation

Fulcrum Digital • Lisboa

Presencial
EUR 55 000 - 75 000
Site Reliability Engineer
Site Reliability Engineer

Fulcrum Digital Inc • Lisboa

Presencial
EUR 50 000 - 70 000
Site Reliability Engineer
Site Reliability Engineer

EPAM Systems • Portugal

Presencial
EUR 40 000 - 60 000
Application Production Support Engineer
Application Production Support Engineer

Inetum • Lisboa

Presencial
EUR 52 000 - 78 000
DevOps Engineer
DevOps Engineer

Fulcrum Digital Inc • Lisboa

Presencial
EUR 42 000 - 62 000
DevOps Engineer
DevOps Engineer

Fulcrum Digital • Lisboa

Presencial
EUR 40 000 - 60 000
Cloud Support Infrastructure Expert
Cloud Support Infrastructure Expert

act digital • Lisboa

Presencial
EUR 42 000 - 66 000
IT Operations Engineer (DevOps & Production Services)
IT Operations Engineer (DevOps & Production Services)

Inetum • Lisboa

Presencial
EUR 42 000 - 64 000