The Senior DevOps Engineer designs, implements and maintains scalable, secure and reliable cloud infrastructure, leading CI/CD, incident response and DevOps standards across the team.
What you'll do
- Design, implement and maintain scalable, secure and reliable cloud infrastructure.
- Develop and optimise CI/CD pipelines to automate software builds, testing and deployments.
- Manage infrastructure using Infrastructure as Code tools such as Terraform, CloudFormation or Bicep.
- Deploy and manage containerised applications using Docker and Kubernetes.
- Establish monitoring, logging and alerting to improve system availability and incident detection.
- Lead troubleshooting, incident response and root cause analysis for infrastructure and deployment issues.
- Implement security controls, access management, secrets management and vulnerability remediation.
- Design and test backup, disaster recovery and high-availability solutions.
- Improve cloud performance, capacity planning and cost efficiency.
- Collaborate with development, testing and security teams to improve release processes.
- Establish DevOps standards, maintain operational documentation and mentor junior engineers.
- Plan infrastructure upgrades and migrations with appropriate testing and rollback procedures.
What we’re looking for
- Bachelor’s degree in Information Technology, Computer Science, Software Engineering or a related discipline, or equivalent professional experience.
- Minimum 5 years of relevant experience, including hands‑on responsibility for production environments.
- Strong expertise in at least one major cloud platform: AWS, Microsoft Azure or Google Cloud.
- Demonstrated experience designing and maintaining CI/CD pipelines using tools such as Jenkins, GitHub Actions, GitLab CI/CD or Azure DevOps.
- Strong Linux administration skills and proficiency in Bash, Python or PowerShell.
- Practical experience with Infrastructure as Code, configuration management and automated provisioning.
- Hands‑on experience with Docker and Kubernetes, including deployment, scaling and troubleshooting.
- Strong understanding of networking, including DNS, TCP/IP, load balancing, firewalls and VPNs.
- Experience with monitoring, logging, incident management and disaster recovery.
- Knowledge of secure deployment practices, identity and access management, and secrets handling.
- Proficiency with Git, branching strategies and release management.
- Strong problem‑solving, technical leadership, communication and mentoring skills.
- Desirable: relevant cloud, DevOps, Kubernetes or Terraform certifications.
- Desirable: experience with hybrid or multi‑cloud infrastructure.
- Desirable: familiarity with GitOps and automated security testing.
- Desirable: knowledge of Site Reliability Engineering practices, including service-level objectives and error budgets.
- Desirable: experience supporting database platforms, microservices or AI/ML workloads.
- Desirable: experience working in Agile teams and coordinating releases across multiple projects.