Get more replies from employers
Send a job-specific resume in minutes.
UKG in Ireland, Leinster is seeking a Senior Engineer to improve service lifecycle from design to operations. Candidates should have over 4 years of experience in engineering roles, especially with AWS in production. Your responsibilities include managing multiple environments, enhancing system reliability, and mentoring team members.
You will also engage in incident response and on-call rotations, focusing on operational excellence through automation and best practices.
Engage in and improve the full lifecycle of services from conception to end-of-life, including system design input, operational readiness, and capacity planning. Define and evolve standards and best practices related to system architecture, service delivery, metrics, and operational automation. Own the management and health of multiple environments, including provisioning, consistency, lifecycle, and operational readiness. Partner with engineering teams to support delivery and operability, providing guidance, tooling, and frameworks that scale beyond one-off requests. Improve system reliability, application delivery, and operational efficiency through automation, post-incident reviews, and continuous learning. Treat operational problems as software engineering problems, with a focus on reducing toil and increasing clarity. Own the operational impact of changes, including observability, rollback readiness, and on-call safety. Actively participate in incident response and on-call rotations. As a senior engineer, lead by influence, mentor others, and contribute to architectural and operational decision-making.
4+ years of hands‑on experience in engineering, cloud, or reliability-focused roles.
Hands‑on experience with AWS in production environments is required. Exposure to GCP is a plus for future platform evolution.
Experience operating or supporting customer‑facing systems at scale.
Ability to write and maintain production-quality code or automation (e.g. Python, JavaScript, Java, Go).
Strong understanding of version control, CI/CD concepts, and system reliability fundamentals.
Experience with containerised or distributed systems.
Experience with modern delivery and observability tooling, such as:
Knowledge of infrastructure, networking, and security fundamentals.
Experience with observability concepts (metrics, logs, tracing).
Comfort working in distributed teams across time zones.
Demonstrated ability to reason about failure modes and operational risk.