Senior Site Reliability Engineer

Acuity

Munster

On-site

EUR 110,000 - 150,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Acuity's Reflect platform in Cork seeks a Senior Site Reliability Engineer to lead reliability, automation, and production ownership across cloud services.

You will work at the intersection of software engineering and cloud infrastructure, building resilient systems, defining SLOs, and improving runbooks while collaborating with global engineers and product teams in a fast-growing SaaS environment.

Qualifications

  • 5+ years of professional experience in software engineering, SRE, or related role.
  • Strong hands-on experience with Microsoft Azure (AKS, Monitor, Key Vault, ACR, VNets).
  • Experience with containerized and orchestrated environments (Docker, Kubernetes).
  • Experience operating production SaaS systems at scale.
  • Excellent communication and collaboration skills.
  • Willingness to stay up-to-date with emerging technologies.

Responsibilities

  • Own the availability, reliability, and performance of Reflect’s production environments and more products to be onboarded.
  • Define, track, and report on service health metrics including uptime, availability, and reliability indicators.
  • Drive root cause analysis (RCA), analyze system logs and ensure corrective and preventative actions are implemented.
  • Part of a global team providing operational and escalation coverage, leading incident response and recovery for critical services.
  • Automate operational workflows to reduce manual toil and improve consistency.
  • Support and improve deployment processes for features, patches, and hotfixes while maintaining a strong security posture.
  • Create, maintain, and continuously improve runbooks and standard operating procedures (SOPs).
  • Design and evolve monitoring, alerting, and observability standards across the platform.
  • Build and maintain dashboards and alerts that provide clear, actionable insight into system health.
  • Ensure monitoring supports SLOs and operational decision-making, not just data collection.
  • Work closely with software engineers to embed reliability best practices into system design and delivery.
  • Partner with product and platform teams to translate business requirements into reliable, scalable technical solutions.
  • Contribute to a culture of shared production ownership and continuous improvement.
  • Stay up-to-date with the latest technologies and industry trends to drive innovation.

Skills

Azure
AKS
Azure Monitor
Key Vault
ACR
Docker
Kubernetes
SRE practices
Incident management
Communication

Education

Bachelor's degree in Computer Science

Tools

Azure DevOps Pipelines
Prometheus
Grafana
Datadog
SQL
NoSQL
Cloud security

Job description

Job Summary

Reflect is Acuity’s flagship cloud platform, providing monitoring, management, and operational intelligence for QSC's Q-SYS Full stack AV platform and connected systems used by enterprise customers worldwide. As Reflect continues to scale as a mission‑critical SaaS service, reliability, availability, and operational maturity are central to both customer trust and long‑term growth.

This role is based in Cork, within Acuity’s Digital Centre of Excellence (DCOE)—a growing hub for cloud, platform, and product engineering. The DCOE plays a central role in defining how Acuity builds, operates, and scales its cloud‑first platforms, working closely with global engineering, product, and operations teams.

We are seeking a Senior Site Reliability Engineer (SRE) to help define and raise the reliability bar for Reflect’s cloud services. You will be a senior, hands‑on contributor, operating at the intersection of software engineering, cloud infrastructure, and production operations. This role is focused on engineering‑led reliability: building resilient systems, automating operational workflows, and embedding reliability practices directly into how Reflect is designed and delivered.

As part of the Cork DCOE, you will have meaningful ownership and influence—shaping observability standards, incident response, SLOs, runbooks, and platform reliability patterns used across Reflect. You will collaborate closely with software engineers, DevOps, and product teams, ensuring Reflect can scale safely while enabling teams to move fast and with confidence.

This is an opportunity to join a growing cloud engineering centre in Cork, work on a globally‑used SaaS platform, and help define what “good” looks like for reliability, operability, and production ownership as Reflect continues to mature as an enterprise‑grade cloud service.

Key Tasks & Responsibilities (Essential Functions)
Reliability & Production Ownership
  • Own the availability, reliability, and performance of Reflect’s production environments and more products to be onboarded
  • Define, track, and report on service health metrics including uptime, availability, and reliability indicators.
  • Drive root cause analysis (RCA), analyze system logs and ensure corrective and preventative actions are implemented.
  • Part of a global team providing operational & escalation coverage, leading incident response and recovery for critical services.
Automation & Operational Excellence
  • Automate operational workflows to reduce manual toil and improve consistency.
  • Support and improve deployment processes for features, patches, and hotfixes while maintaining a strong security posture.
  • Create, maintain, and continuously improve runbooks and standard operating procedures (SOPs).
Observability & Monitoring
  • Design and evolve monitoring, alerting, and observability standards across the platform.
  • Build and maintain dashboards and alerts that provide clear, actionable insight into system health.
  • Ensure monitoring supports SLOs and operational decision‑making, not just data collection.
Collaboration & Engineering Enablement
  • Work closely with software engineers to embed reliability best practices into system design and delivery.
  • Partner with product and platform teams to translate business requirements into reliable, scalable technical solutions.
  • Contribute to a culture of shared production ownership and continuous improvement.
  • Stay up‑to‑date with the latest technologies and industry trends to drive innovation.
Skills and minimum Expertise
  • 5+ years of professional experience in software engineering, SRE, or a related role.
  • Strong hands‑on experience with Microsoft Azure, including services such as:
  • AKS, Azure Monitor / Log Analytics, Key Vault, ACR, VNets, Managed Identity.
  • Deep experience with containerized and orchestrated environments (Docker, Kubernetes).
  • Proven experience operating and supporting production SaaS systems at scale.
  • A keen eye for detail and a knack for troubleshooting complex issues.
  • Excellent communication and collaboration skills, with the ability to work effectively across teams.
  • A passion for learning and a drive to stay up‑to‑date with emerging technologies.
Preferred Skills and Experience
  • Experience with CI/CD pipelines and automation, preferably Azure DevOps Pipelines.
  • Familiarity with modern monitoring and alerting tools (e.g. Prometheus, Grafana, Datadog).
  • Knowledge of SQL and NoSQL databases.
  • Understanding of cloud security best practices.
  • Exposure to DevOps and SRE principles and practices to streamline development and deployment processes.
  • Azure and/or Kubernetes certifications.
Education
  • Bachelors degree in Computer Science or a related field.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Acuity Brands, Inc. • Cork

Hybrid
EUR 90,000 - 130,000
Senior Cloud SRE: Reliability, Observability & Automation
Senior Cloud SRE: Reliability, Observability & Automation

Acuity • Munster

On-site
EUR 110,000 - 150,000
Senior Cloud SRE: Reliability, Observability & Automation
Senior Cloud SRE: Reliability, Observability & Automation

Acuity Brands, Inc. • Cork

Hybrid
EUR 90,000 - 130,000
Development & Product Management Site Reliability Engineering Technical Lead Dublin, Ireland
Development & Product Management Site Reliability Engineering Technical Lead Dublin, Ireland

AMCS Group • Dublin

On-site
EUR 90,000 - 130,000
Development & Product Management Site Reliability Engineering Technical Lead Galway, Ireland
Development & Product Management Site Reliability Engineering Technical Lead Galway, Ireland

AMCS Group • Galway

On-site
EUR 100,000 - 140,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Dublin

On-site
EUR 110,000 - 150,000
Development & Product Management Site Reliability Engineering Technical Lead Limerick, Ireland
Development & Product Management Site Reliability Engineering Technical Lead Limerick, Ireland

AMCS Group • Limerick

On-site
EUR 75,000 - 95,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Limerick

On-site
EUR 110,000 - 140,000
Cloud Infrastructure/Engineer
Cloud Infrastructure/Engineer

Mars Capital • Dublin

On-site
EUR 80,000 - 105,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobtailor • Ireland

On-site
EUR 70,000 - 120,000