Sr System Reliability Engineer (Application Support + Automation)

fulcrumdigital

Dublin

On-site

EUR 70,000 - 110,000

Full time

11 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Fulcrum Digital is seeking a Production/DevOps engineer to plan, manage and oversee all aspects of the production environment. You will define monitoring strategies, respond to incidents and drive measures to reduce incident frequency across multiple environments.

You will support CI/CD deployment, automate processes, and contribute to an end-to-end lifecycle of services from design to operation. On-call duties and collaboration with a global team are required.

Qualifications

  • Must have strong Linux and shell scripting skills.
  • Experience with ITIL/ITSM processes and incident management.
  • Experience with monitoring, alerting, and performance optimization.
  • Familiarity with CI/CD pipelines and deployment automation.

Responsibilities

  • Plan, manage, and oversee production environment activities.
  • Define strategies for application performance monitoring and optimization in prod.
  • Respond to incidents and improve platform reliability and MTTR.
  • Support CI/CD pipelines and promote code across environments.
  • Automate processes and contribute to DevOps best practices.

Skills

Linux
Shell scripting
ITIL/ITSM
Incident response
Problem solving

Tools

Splunk
Dynatrace
Jenkins CI/CD
Groovy
Yaml
Git / Bitbucket
Ansible
Chef

Job description

Who are we

Fulcrum Digital is an agile and next-generation digital accelerating company providing digital transformation and technology services right from ideation to implementation. These services have applicability across a variety of industries, including banking & financial services, insurance, retail, higher education, food, healthcare, and manufacturing.


The Role

Plan, manage, and oversee all aspects of a Production Environment Define strategies for Application Performance Monitoring, Optimization in Prod environment Respond to Incidents and improvise platform based on feedback and measure the reduction of incidents over time. Support deployment of code into multiple lower environments. Supporting current processes with an emphasis on automating everything as soon as possible. Design, develop and standardize Monitoring and Alerting mechanism for the supported applications. Take a holistic approach to problem solving, by connecting the dots during a production event through the various technology stack that makes up the platform, to optimize meantime to recover. Engage in and improve the whole lifecycle of services—from inception and design, through deployment, operation and refinement. Analyze ITSM activities of the platform and provide feedback loop to development teams on operational gaps or resiliency concerns. Support services before they go live through activities such as system design consulting, capacity planning and launch reviews. Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead in DevOps automation and best practices. Maintain services once they are live by measuring and monitoring availability, latency and overall system health. Scale systems sustainably through mechanisms like automation and evolving systems by pushing for changes that improve reliability and velocity. Work with a global team spread across tech hubs in multiple geographies and time zones. Ability to share knowledge and explain processes and procedures to others. Share knowledge and mentor junior resources Able to perform on-call duties on a rotational basis. Occasional off hours work required.


Requirements

Skills – Must Have Linux Mainframe Shell Scripting ITIL / ITSM, Application Troubleshooting SQL Any Monitoring tool (Preferred Splunk/Dynatrace) Jenkins - CI/CD Groovy Scripting/Yaml - basic Git basic/bit bucket - basic Ansible/Chef - good to have Good To Have Even Framework architecture

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr System Reliability Engineer (Application Support + Automation)
Sr System Reliability Engineer (Application Support + Automation)

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
Senior SRE: Automation, Monitoring & Incident Response
Senior SRE: Automation, Monitoring & Incident Response

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
Senior Platform Reliability Engineer
Senior Platform Reliability Engineer

fulcrumdigital • Dublin

On-site
EUR 70,000 - 110,000
Site Reliability Engineer - Tier 1
Site Reliability Engineer - Tier 1

Xanadu Consultancy • Ireland

On-site
EUR 35,000 - 55,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Dublin

On-site
EUR 110,000 - 150,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Leinster

On-site
EUR 90,000 - 130,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobtailor • Ireland

On-site
EUR 70,000 - 120,000
Site Reliability Engineer (SRE) / DevOps Engineer
Site Reliability Engineer (SRE) / DevOps Engineer

Fulcrum Digital • Dublin

On-site
EUR 85,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

Back4good • Ireland

On-site
EUR 60,000 - 80,000
Strong and attractive package
Relocation assistance
Development & Product Management Site Reliability Engineering Technical Lead Dublin, Ireland
Development & Product Management Site Reliability Engineering Technical Lead Dublin, Ireland

AMCS Group • Dublin

On-site
EUR 90,000 - 130,000