System Reliability Engineer (Application Support + Automation)

Fulcrum Digital Inc

Lisboa

Presencial

EUR 42 000 - 65 000

Tempo integral

Há 5 dias
Torna-te num dos primeiros candidatos

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Fulcrum Digital is seeking a Production Environment Engineer to plan, manage, and oversee all aspects of the production stack. You will define strategies for application performance monitoring and optimization in the production environment.

You will respond to incidents and drive improvements based on feedback, supporting deployments across multiple environments and contributing to automation efforts to accelerate delivery and reliability.

Qualificações

  • Operates and maintains production environments with focus on reliability and performance.
  • Defines and implements monitoring, alerting, and incident response processes.
  • Supports CI/CD pipelines and automated deployments across environments.
  • Engages in capacity planning, system design reviews, and production readiness activities.

Responsabilidades

  • Plan, manage and oversee all aspects of a Production Environment.
  • Define strategies for Application Performance Monitoring and Prod optimization.
  • Respond to incidents and improve platform based on feedback.
  • Support deployment of code into multiple lower environments and automate processes.

Conhecimentos

ITIL / ITSM
DevOps
Incident response
Capacity planning

Ferramentas

Linux
Shell Scripting
SQL
Splunk
Dynatrace
Jenkins
Ansible
Groovy Scripting
Yaml
Git

Descrição da oferta de emprego

Who are we

Fulcrum Digital is an agile and next-generation digital accelerating company providing digital transformation and technology services right from ideation to implementation. These services have applicability across a variety of industries, including banking & financial services, insurance, retail, higher education, food, healthcare, and manufacturing.

The Role
  • Plan, manage, and oversee all aspects of a Production Environment
  • Define strategies for Application Performance Monitoring, Optimization in Prod environment
  • Respond to Incidents and improvise platform based on feedback and measure the reduction of incidents over time.
  • Support deployment of code into multiple lower environments. Supporting current processes with an emphasis on automating everything as soon as possible.
  • Design, develop and standardize Monitoring and Alerting mechanism for the supported applications.
  • Take a holistic approach to problem solving, by connecting the dots during a production event through the various technology stack that makes up the platform, to optimize meantime to recover.
  • Engage in and improve the whole lifecycle of services—from inception and design, through deployment, operation and refinement.
  • Analyse ITSM activities of the platform and provide feedback loop to development teams on operational gaps or resiliency concerns.
  • Support services before they go live through activities such as system design consulting, capacity planning and launch reviews.
  • Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead in DevOps automation and best practices.
  • Maintain services once they are live by measuring and monitoring availability, latency, and overall system health.
  • Scale systems sustainably through mechanisms like automation and evolving systems by pushing for changes that improve reliability and velocity.
  • Work with a global team spread across tech hubs in multiple geographies and time zones.
  • Ability to share knowledge and explain processes and procedures to others.
  • Able to perform on-call duties on a rotational basis.
  • Occasional off hours work required.
Requirements
  • Linux
  • Shell Scripting
  • ITIL / ITSM
  • SQL - good to have
  • Application Troubleshooting
  • Any Monitoring tool (Preferred Splunk/Dynatrace)
  • Jenkins - CI/CD
  • Ansible
  • Groovy Scripting/Yaml
  • Git basic/bit bucket
Good To Have
  • Payments Flows, Switching, Settlements, Authorisation flows.
  • Even Framework architecture
Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

System Reliability Engineer (Application Support + Automation)
System Reliability Engineer (Application Support + Automation)

Fulcrum Digital • Lisboa

Presencial
EUR 55 000 - 75 000
Site Reliability Engineer
Site Reliability Engineer

Fulcrum Digital Inc • Portugal

Presencial
EUR 50 000 - 80 000
System Reliability Engineer — App Support & Automation
System Reliability Engineer — App Support & Automation

Fulcrum Digital • Lisboa

Presencial
EUR 55 000 - 75 000
Site Reliability Engineer
Site Reliability Engineer

Fulcrum Digital Inc • Lisboa

Presencial
EUR 50 000 - 70 000
DevOps Engineer
DevOps Engineer

Fulcrum Digital • Lisboa

Presencial
EUR 40 000 - 60 000
DevOps Engineer
DevOps Engineer

Fulcrum Digital Inc • Lisboa

Presencial
EUR 42 000 - 62 000
Application Production Support Engineer
Application Production Support Engineer

Inetum • Lisboa

Presencial
EUR 52 000 - 78 000
Site Reliability Engineer
Site Reliability Engineer

EPAM Systems • Portugal

Presencial
EUR 40 000 - 60 000
Cloud Support Infrastructure Expert
Cloud Support Infrastructure Expert

act digital • Lisboa

Presencial
EUR 42 000 - 66 000
Production Reliability Engineer — App Support & Automation
Production Reliability Engineer — App Support & Automation

Fulcrum Digital Inc • Lisboa

Presencial
EUR 42 000 - 65 000