SRE (Application Support + Dev-Ops + Automation)

fulcrumdigital

Dublin

On-site

EUR 70,000 - 110,000

Full time

11 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Fulcrum Digital is seeking a Site Reliability Engineer to enhance observability, automate deployments, and ensure reliable cloud-based services. You will work with Linux/Unix systems, scripting languages, and modern DevOps practices to improve incident response and system health.

Collaborate with engineering teams to deliver scalable and resilient technology solutions, contributing to best practices and documentation.

Qualifications

  • Observability: use scripting/tools to collect metrics, logs, traces and enable incident detection.
  • Programming: write/maintain code and scripts to automate tasks and deployment.
  • Systems Administration: configure and troubleshoot Linux/Unix, networks and security.
  • Cloud: design/deploy/manage apps and infra on cloud platforms (AWS, Azure, GCP).
  • Reliability: design/operate systems for high availability and scalability.
  • DevOps: apply CI/CD, containers, orchestration for faster/reliable delivery.
  • Troubleshooting: diagnose and resolve issues across systems and networks.
  • Capacity Planning: monitor usage and forecast future needs.
  • IT Service Management: incident/problem/change management alignment.
  • Proactive Monitoring: use signals to anticipate issues and drive improvements.

Responsibilities

  • Independently execute key elements of projects within the SRE area, applying in-depth knowledge to resolve roadblocks.
  • Assist in evaluating operational requirements and developing technical solutions within existing frameworks.
  • Support automation and scripting efforts to improve operational workflows and incident response.
  • Troubleshoot and resolve routine and some complex system issues, escalating when necessary to maintain health.
  • Contribute to documentation, knowledge sharing, and best practices to enhance team procedures.
  • Collaborate with development teams and stakeholders to ensure reliability solutions align with business needs.
  • Participate in reviews and quality assurance to uphold system stability.
  • May contribute to solution development for new products/services and manage smaller initiatives as an experienced contributor.

Skills

Observability
Programming & Scripting
Systems Administration
Cloud Computing
Reliability & Scalability
DevOps Practices
Troubleshooting
Capacity Planning
IT Service Management
Proactive Monitoring

Job description

Who are we Fulcrum Digital is an agile and next-generation digital accelerating company providing digital transformation and technology services right from ideation to implementation. These services have applicability across a variety of industries, including banking & financial services, insurance, retail, higher education, food, healthcare, and manufacturing.

Requirements

As part of the Business Operations team, you will:

  • Independently execute key elements of projects/processes within the Site Reliability Engineering area by applying in-depth knowledge of their discipline and area best practices to effectively resolve problems and roadblocks as they occur.
  • Assist in evaluating operational requirements and developing technical solutions within existing frameworks.
  • Support automation and scripting efforts to improve operational workflows and incident response processes.
  • Troubleshoot and resolve routine and some complex system issues, escalating when necessary to maintain system health.
  • Contribute to documentation, knowledge sharing, and best practices to enhance team operational procedures.
  • Collaborate with development teams and stakeholders to ensure reliability solutions align with technical and business needs.
  • Participate in reviews and quality assurance activities to uphold system stability standards.
  • May contribute to solution development for new products/services and/or manage smaller project/initiatives as an experienced individual contributor with specialized knowledge within the Site Reliability Engineering area.
Role qualifications
  • Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement.
  • Programming and Scripting - Ability to write and maintain code and scripts to automate tasks, build operational tools, and support monitoring, deployment, and incident response using languages such as Python, Go, Bash, or similar.
  • Systems and Network Administration - Ability to configure, operate, and troubleshoot Linux/Unix systems and network components, applying knowledge of networking concepts, protocols, security, and system reliability.
  • Cloud Computing and Infrastructure - Ability to design, deploy, and manage applications and infrastructure on cloud platforms (e.g., AWS, Azure, GCP), ensuring scalability, security, availability, and operational efficiency.
  • Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand
  • DevOps Practices - Ability to apply DevOps principles and practices, including CI/CD pipelines, containerization, and orchestration, to enable faster, more reliable software delivery and operations.
  • Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations.
  • Capacity Planning and Performance Optimization - Ability to monitor resource utilization, forecast future capacity needs, and optimize system performance to support growth, scalability, and efficient infrastructure usage.
  • IT Service Management - Ability to apply IT service management principles to incident, problem, and change management, ensuring reliable service delivery, effective incident response, and continuous service improvement aligned to business needs.
  • Proactive Monitoring and Improvement (SRE Applications) - The ability to use application reliability signals to anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE (Application Support + Dev-Ops + Automation)
SRE (Application Support + Dev-Ops + Automation)

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
SRE (Application Support + Dev-Ops + Automation)
SRE (Application Support + Dev-Ops + Automation)

Fulcrum Digital Inc • Dublin

On-site
EUR 90,000 - 120,000
Senior Site Reliability Engineer (SRE) – Business Operations
Senior Site Reliability Engineer (SRE) – Business Operations

Moofwd • Dublin

Hybrid
EUR 110,000 - 150,000
Sr System Reliability Engineer (Application Support + Automation)
Sr System Reliability Engineer (Application Support + Automation)

Fulcrum Digital Inc • Cavan

On-site
EUR 70,000 - 100,000
Sr System Reliability Engineer (Application Support + Automation)
Sr System Reliability Engineer (Application Support + Automation)

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
Senior Production SRE - Reliability & Automation
Senior Production SRE - Reliability & Automation

Fulcrum Digital Inc • Cavan

On-site
EUR 70,000 - 100,000
SRE: Reliability & Automation Engineer
SRE: Reliability & Automation Engineer

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Harvey Nash • Dublin

On-site
EUR 90,000 - 130,000
Senior SRE: Automation, Monitoring & Incident Response
Senior SRE: Automation, Monitoring & Incident Response

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
Senior SRE: Reliability & Automation Architect
Senior SRE: Reliability & Automation Architect

Fulcrum Digital Inc • Dublin

On-site
EUR 90,000 - 120,000