Network Production Engineer, Network.AI

Meta

Greater London

On-site

GBP 120,000 - 180,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Meta's Network Infrastructure team seeks a seasoned Network Production Engineer in the UK to design, build, and operate our global backbone, data center, and AI networks. You will create scalable networking solutions, develop automation, and collaborate across hardware and software teams to ensure reliability and efficiency.

You will work with multi-vendor devices, routing protocols, and distributed systems at scale, contributing to cutting-edge AI workloads and network innovations.

Qualifications

  • Bachelor's degree in Computer Science, Computer Engineering, or equivalent practical experience.
  • 6+ years designing, building, operating scalable systems and networks.
  • Experience with multiple programming languages and rapid language learning.
  • Knowledge of TCP, IPv4/6, BGP/MPLS/ISIS and related services.
  • Experience configuring multi-vendor network devices.

Responsibilities

  • Establish and implement global best practices and design scalable network solutions.
  • Develop automation and tooling to support deployment and operations.
  • Design solutions across hardware platforms and networks.
  • Lead automation enhancements for CI, testing, release, and config management.
  • Investigate complex technical issues across networks and tooling.
  • Drive process improvements and cross-team collaboration.
  • Participate in on-call rotations and mentor teammates.

Skills

Programming languages
Network protocols
BGP/MPLS/ISIS
Python
C/C++/Go
Networking design

Education

Bachelor's degree in CS/Engineering
Master's degree or grad work

Tools

Juniper
Cisco
Arista
Brocade

Job description

About

The Network Infrastructure team is responsible for designing, building, and operating one of the largest networks in the world. Networking is at the core of all Meta products and experiences, and we are looking for Network Production Engineers who are interested in solving complex technical challenges in the Backbone, Datacenter, and AI Network domains.The scale of the network and its continuous expansion presents an opportunity to work on and solve interesting engineering challenges in the datacenter network domain. We create new and innovative ways of designing and operating our global datacenter networks and do it at scale with efficiency.Production Network Engineers at Meta are hybrid software and network engineers who design, build, and operate our worldwide network. This team owns the complete lifecycle of the network, which includes areas of planning, design, product definition, QA, deployment, and monitoring. Simple, elegant, and scalable network design, automation, and data analytics are the keys to meeting our demands. In this role, you will be part of a team that is responsible for conceiving design solutions, developing and deploying network software, systems, and tools that keep the network operating at maximum reliability, scalability, and efficiency.This role offers an opportunity to solve scaling challenges supporting billions of people using our family of apps, as well as cutting-edge challenges in AI workloads that power new Meta products.

Responsibilities
  • Establish and implement global best practices and design new scalable network solutions
  • Conceptualize, build, and maintain automation and tools to support new product introductions, network deployment, release engineering, and operations
  • Design and develop solutions that scale across a variety of hardware platforms of network equipment
  • Lead enhancements of automation for continuous integration, validations, testing infrastructure, release, and configuration management across our global backbone, data center, and edge networks
  • Work closely with our hardware, software, and sourcing teams to develop new networking solutions and influence the future of networking and its associated infrastructure
  • Conduct thorough investigations into complex technical issues across networks, ranging from automated tooling to hardware failures and network issues
  • Develop operational process improvements and implement them in scalable, automated workflows to enhance operational efficiency
  • Help increase operational efficiency between peers and cross-functional teams by identifying roadblocks, designing and delivering automation solutions, and driving change
  • Proactively identify gaps that impact multiple teams, develop execution plans, drive projects to completion, and influence other teams to align on solutions
  • Participate in an on-call rotation to learn from real-world production challenges and take the lessons to improve current and future generation products
  • Contribute to team growth and development through peer mentorship
Minimum Qualifications
  • Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
  • 6+ years of experience planning, designing, building, and/or operating scalable and reliable systems and/or networks
  • Experience coding in at least one programming language (e.g., Python, C++, Go, Rust, or Java) and rapidly learning new development languages, software, frameworks, and APIs
  • Demonstrated knowledge of TCP, IPv4/6, Routing Protocols (one or more of BGP, MPLS, ISIS, or similar), and related network services (e.g., DHCP and DNS)
  • Experience developing and understanding network device configuration in multi-vendor environments (Juniper, Cisco, Arista, Brocade, etc.)
  • Experience in configuration and maintenance of network devices and NMS systems, or applications such as web servers, load balancers, relational databases, storage systems, and messaging systems
  • Hardware evaluation and vendor management experience Proven experience designing, developing, and operating distributed systems at scale, with an in-depth understanding of the challenges and opportunities in this space
  • Experience designing and maintaining automated testing infrastructure to ensure the quality and reliability of network systems and tooling
  • Master's degree or graduate work experience in Computer Science, Computer Engineering, or a related technical field
  • Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
  • In-depth knowledge of software and network debugging, profiling, and instrumentation techniques to ensure optimal system performance
  • Knowledge of IB/RDMA/RoCE Networks, including RDMA congestion control mechanisms, AI training workloads and demands they exert on networks
  • Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
  • 6+ years of experience building software solutions for managing network infrastructure, with a focus on scalability and reliability
  • Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Network Production Engineer, Global Backbone & AI Networking
Network Production Engineer, Global Backbone & AI Networking

Meta • Greater London

On-site
GBP 120,000 - 180,000
Production Engineering Manager, Rotational Network Engineering (RNE) Program
Production Engineering Manager, Rotational Network Engineering (RNE) Program

Meta • Greater London

On-site
GBP 90,000 - 130,000
Production Engineering Manager
Production Engineering Manager

Meta • Greater London

Hybrid
GBP 120,000 - 170,000
Network Engineer
Network Engineer

OpenAI • Greater London

Hybrid
GBP 85,000 - 125,000
Private medical insurance
Pension plan
Maternity leave 52 weeks
+5
Staff Network Engineer
Staff Network Engineer

Nscale • Greater London

On-site
GBP 90,000 - 130,000
Principal Network Engineer
Principal Network Engineer

Nscale • Greater London

On-site
GBP 120,000 - 170,000
Base + equity
Flexible working
Competitive package
+1
Principal Network Engineer
Principal Network Engineer

Jobtailor • Greater London

On-site
GBP 110,000 - 170,000
Sr Network Engineer, Reliability and Observability
Sr Network Engineer, Reliability and Observability

Blue Signal Search • Greater London

On-site
GBP 90,000 - 140,000
Equity participation
Comprehensive benefits package
Network Engineer
Network Engineer

Selby Jennings • Greater London

On-site
GBP 95,000 - 140,000
HPC Network Engineer
HPC Network Engineer

Fuse Energy, LLC • Greater London

Hybrid
GBP 85,000 - 120,000
Fully expensed tech to match your need
Private health insurance
Breakfast and dinner allowance