Site Reliability Engineer, Dublin

Apple Inc.

Dublin

On-site

EUR 110,000 - 140,000

Full time

19 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Apple Inc. in Dublin is seeking a Site Reliability Engineer to design, build, and operate scalable compute platforms including VMs and containers. You will own reliability, observability, and automation across mission-critical services for global users.

You will collaborate with software, QA, and program management to embed reliability into development and deployment lifecycles, while contributing to architecture and disaster recovery exercises.

Qualifications

  • Expert and in-depth experience with cloud operations focused on infrastructure-as-a-service (compute, storage, network virtualization).
  • Strong software development skills in Go and Java, with experience building production services, tools or automation frameworks.
  • Experience with software development lifecycle practices including version control, code review, CI/CD, and automated testing.
  • Experience operating and engineering large-scale multi-tenant Infrastructure as a Managed service.
  • Ability to articulate complex technical concepts to both technical and non-technical stakeholders.

Responsibilities

  • Design and develop tooling, frameworks, and automation in Go and Java to improve reliability, scalability, and operational efficiency of compute infrastructure (VMs, containers, orchestration).
  • Define and implement SLOs/SLIs for compute services and build the observability pipelines (metrics, logging, tracing) to measure and enforce them.
  • Lead incident response for compute infrastructure, driving triage, root cause analysis, and postmortem corrective actions.
  • Develop and maintain infrastructure-as-code and CI/CD pipelines, ensuring reproducibility, automated testing, and staged rollouts across the fleet.
  • Contribute to compute platform architecture through design reviews, technical design documents, production readiness reviews, capacity planning, and disaster recovery exercises.
  • Partner cross-functionally with engineering, QA, and program management to embed reliability into the development lifecycle, upholding best practices in code review, testing, and documentation.

Skills

Go
Java
Production services
Automation frameworks
CI/CD
Code review
Version control
Automated testing
Multi-tenant Infra
Communication

Tools

OpenStack
CloudStack
Libvirt
QEMU
KVM
Splunk
Grafana
Prometheus

Job description

Selection changes the language of the page/content

Site Reliability Engineer, Dublin

Dublin, County Dublin, Ireland Software and Services

People at Apple don’t just build products — they craft the kind of experience that has revolutionised entire industries. The diverse collection of our people and their ideas inspire innovation in everything we do. Imagine what you could do here! Join Apple, and help us leave the world better than we found it.The Apple Services Engineering(ASE) team builds and provides systems and infrastructure that power Apple’s services (such as iCloud, Apple Music, Apple Intelligence, Maps and more). We are the foundation on which Apple’s software developers build the products that our customers love.Our services have to scale globally, stay highly available, and “just work.” If you love designing, engineering and running systems and infrastructure that will help millions of customers, then this is the place for you!

Description

Apple Service Engineering (ASE)’s Compute team is seeking highly motivated software engineer with strong technical and communication skills to join our SRE team on our quest to build and enhance massive clusters hosting Virtual Machines, Containers and associated infrastructure that can scale to meet the demands of Apple’s Services offerings. You will work with world-class engineers on core components of Virtualization and Containerization technologies, customize it to help fit Apple’s diverse needs, and engage with the upstream community to drive Apple’s requirements.Ultimately, you will help build the platform that delivers our applications at scale to our end users. As a Compute Site Reliability Engineer, you will be part of the team responsible for providing the platform for mission-critical cloud systems to maintain constant uptime, scale seamlessly, and allow for new applications and services to flourish.

Responsibilities
  • Design and develop tooling, frameworks, and automation in Go and Java to improve reliability, scalability, and operational efficiency of compute infrastructure (VMs, containers, orchestration).
  • Define and implement SLOs/SLIs for compute services and build the observability pipelines (metrics, logging, tracing) to measure and enforce them.
  • Lead incident response for compute infrastructure, driving triage, root cause analysis, and postmortem corrective actions.
  • Develop and maintain infrastructure-as-code and CI/CD pipelines, ensuring reproducibility, automated testing, and staged rollouts across the fleet.
  • Contribute to compute platform architecture through design reviews, technical design documents, production readiness reviews, capacity planning, and disaster recovery exercises.
  • Partner cross-functionally with engineering, QA, and program management to embed reliability into the development lifecycle, upholding best practices in code review, testing, and documentation.
Minimum Qualifications
  • Must be an expert and have in-depth professional experience with cloud operations, with a focus on “infrastructure-as-a-service” (compute, storage, and network virtualization).
  • Strong software development skills in Go and Java, with experience building production services, tools or automation frameworks.
  • Experience with software development lifecycle practices including version control, code review, CI/CD, and automated testing.
  • Experience operating and engineering large-scale multi-tenant Infrastructure as a Managed service
  • Ability to articulate complex technical concepts to both technical and non-technical stakeholders.
Preferred Qualifications
  • Experience with Infrastructure as a Service orchestration tools (OpenStack, CloudStack, etc) is a plus
  • Experience with Linux system virtualization (Libvirt, QEMU, KVM, etc), along with the APIs
  • Ability to implement and coordinate telemetry using monitoring and observability tools such as Splunk, Grafana, and Prometheus
  • Experience building internal platforms or developer tooling and familiarity with distributed systems concepts

At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.
Learn about accessibility in Apple’s workplace

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineering Manager (SRE) - Apple Services Engineering, Dublin
Software Engineering Manager (SRE) - Apple Services Engineering, Dublin

Apple • Dublin

On-site
EUR 150,000 - 180,000
Site Reliability Engineering Manager (Software) - Apple Services Engineering, Dublin
Site Reliability Engineering Manager (Software) - Apple Services Engineering, Dublin

Apple Inc. • Dublin

On-site
EUR 150,000 - 200,000
Site Reliability Engineering Manager (Software) - Apple Services Engineering, Dublin
Site Reliability Engineering Manager (Software) - Apple Services Engineering, Dublin

Lex • Dublin

On-site
EUR 140,000 - 195,000
Site Reliability Engineer
Site Reliability Engineer

Apple Inc. • Dublin

On-site
EUR 110,000 - 180,000
Site Reliability Engineer, Enterprise Technology Services
Site Reliability Engineer, Enterprise Technology Services

Apple Inc. • Cork

On-site
EUR 90,000 - 130,000
Site Reliability Engineer – Scalable Cloud Platforms
Site Reliability Engineer – Scalable Cloud Platforms

Apple Inc. • Dublin

On-site
EUR 110,000 - 140,000
Cloud SRE: Scale & Reliability for Global Services
Cloud SRE: Scale & Reliability for Global Services

Apple Inc. • Dublin

On-site
EUR 110,000 - 180,000
Cloud SRE Engineer - Scale Global Services
Cloud SRE Engineer - Scale Global Services

Apple • Dublin

On-site
EUR 110,000 - 145,000
SRE Manager: Cloud Platform & Scale Leader
SRE Manager: Cloud Platform & Scale Leader

Apple Inc. • Dublin

On-site
EUR 150,000 - 200,000
Senior Full Stack Engineer, Employee Engagement Engineering
Senior Full Stack Engineer, Employee Engagement Engineering

Apple • Cork

On-site
EUR 90,000 - 150,000