We are looking for a Manager, Infrastructure Engineering to lead our Network and Datacentre Services function.
Your primary mission is to build a high-performing engineering team and deliver secure, resilient, and scalable network services across our on-premises datacentres and hybrid cloud environment. You will set the technical direction, improve engineering practices, and ensure our infrastructure can support the organisation's current needs and future growth.
This is an engineering leadership role with meaningful technical involvement. Your primary focus will be people leadership, strategy, technical governance, and execution. You will also remain closely involved in critical architecture decisions, major incidents, reliability reviews, and complex engineering challenges.
You will work with infrastructure, cloud, security, application, and business teams to improve reliability, accelerate delivery, and reduce operational risk.
What You Will Do:
Lead People and Engineering Execution
- Lead, coach, and develop a team of network engineers.
- Set clear priorities, establish measurable goals, and create accountability for engineering outcomes.
- Build a culture of technical excellence, collaboration, ownership, and continuous learning.
- Support career development, succession planning, and capability building across the team.
- Plan and deliver network and datacentre initiatives in partnership with infrastructure, cloud, security, and application teams.
Set the Technical Direction
- Define and execute the roadmap for network and datacentre services across enterprise datacentres and hybrid cloud platforms.
- Act as the trusted technical leader for network architecture, routing, switching, connectivity, security, capacity, and resiliency.
- Establish engineering standards, reference architectures, lifecycle plans, and operational practices.
- Lead architecture reviews and guide decisions involving scalability, performance, security, reliability, and cost.
- Translate business and technology requirements into practical infrastructure strategies.
- Evaluate new technologies based on operational value, security, maintainability, and long-term business needs.
Improve Reliability and Resilience
- Apply Site Reliability Engineering principles to improve network availability, performance, and operational efficiency.
- Define and track service-level objectives and other reliability measures for critical network services.
- Strengthen monitoring, observability, alerting, and performance-management capabilities.
- Lead or support the response to major network incidents and ensure timely communication with stakeholders.
- Facilitate root cause analysis and ensure corrective actions address underlying technical and process issues.
- Oversee capacity planning, failover testing, resilience reviews, and disaster recovery validation.
- Partner with service owners to ensure network services meet availability, security, compliance, and recovery objectives, including applicable RTO and RPO targets.
Scale Through Automation
- Drive the adoption of Infrastructure-as-Code and network automation practices.
- Use Terraform, Ansible, APIs, and Git-based workflows to standardise provisioning, configuration, compliance, and operational processes.
- Identify repetitive or high-risk manual activities and replace them with reliable, testable automation.
- Promote version control, peer review, reusable modules, and automated validation across the engineering lifecycle.
- Improve delivery speed and consistency while reducing configuration drift and operational risk.
- Provide engineering leadership for secure network architecture, segmentation, connectivity, and access.
- Partner with cybersecurity teams to design and improve firewalls, web application firewalls, DDoS protection, secure access solutions, and related controls.
- Ensure network designs align with applicable security frameworks, hardening standards, and compliance requirements.
- Prioritise vulnerability remediation and help teams address network-related security risks.
- Embed security considerations into architecture, automation, operational processes, and lifecycle planning.
What You Need to Succeed:
- A bachelor's degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience.
- Seasoned expertise in enterprise network, infrastructure, or datacentre engineering.
- Experience leading, coaching, or developing infrastructure or network engineering teams.
- Strong knowledge of enterprise routing, switching, connectivity, network architecture, and operational support.
- Experience designing or operating infrastructure across enterprise datacentres and hybrid environments.
- Demonstrated ownership of infrastructure reliability, with an understanding of dependencies across network, compute, and storage services.
- Hands-on experience with Terraform, Ansible, or comparable infrastructure automation technologies.
- Experience making technical decisions for secure, scalable, and highly available infrastructure.
- Knowledge of monitoring, observability, performance management, incident response, and root cause analysis.
- Strong written and verbal communication skills, with the ability to explain complex technical decisions to technical and non-technical stakeholders.
- Sound judgement and a practical approach to balancing reliability, security, delivery speed, and operational risk.
Preferred Experience
- Experience applying Site Reliability Engineering practices to infrastructure or network services.
- Experience with public cloud networking and hybrid connectivity.
- Knowledge of network security technologies, including firewalls, WAF, DDoS protection, network segmentation, and secure access architectures.
- Experience with APIs, GitOps, continuous integration, or automated infrastructure testing.
- Experience leading disaster recovery exercises, failover testing, capacity planning, or resilience programmes.
- Familiarity with industry security frameworks, compliance requirements, and infrastructure hardening standards.
- Advanced certifications in networking, cloud, security, automation, or infrastructure technologies.