Network Production Engineer, Network.AI

Socket.dev

Walla Walla (WA)

On-site

USD 154,000 - 217,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Bonus
Equity
Benefits

Job summary

Meta is seeking Network Production Engineers to design, build, and operate our massive global backbone, data center, and edge networks. You will own the complete lifecycle from planning to monitoring, and contribute to scalable, automated networking and AI-driven workloads across our platforms.

Ideal candidates combine software and networking expertise, with hands-on experience across multi-vendor devices and modern data center environments.

Qualifications

  • 6+ years of scalable network design/operation experience.
  • Experience coding in Python, C++, Go, Rust, or Java.
  • Knowledge of TCP, IPv4/IPv6, BGP/ISIS/MPLS and DNS/DHCP.

Responsibilities

  • Establish and implement global best practices and scalable network solutions.
  • Build and maintain automation/tools for product introductions and operations.
  • Design solutions that scale across hardware platforms of network equipment.
  • Lead automation enhancements for CI, testing, deployment, and configuration.
  • Collaborate with hardware, software, and sourcing teams on new networking solutions.
  • Investigate complex network issues spanning tooling and hardware failures.
  • Improve operational efficiency with scalable, automated workflows.

Skills

Programming languages
Networking knowledge
Multi-vendor devices

Education

Bachelor's in CS/Engineering
Master's preferred

Tools

Juniper
Cisco
Arista
Brocade

Job description

The Network Infrastructure team is responsible for designing, building, and operating one of the largest networks in the world. Networking is at the core of all Meta products and experiences, and we are looking for Network Production Engineers who are interested in solving complex technical challenges in the Backbone, Datacenter, and AI Network domains. The scale of the network and its continuous expansion presents an opportunity to work on and solve interesting engineering challenges in the datacenter network domain. We create new and innovative ways of designing and operating our global datacenter networks and do it at scale with efficiency. Production Network Engineers at Meta are hybrid software and network engineers who design, build, and operate our worldwide network. This team owns the complete lifecycle of the network, which includes areas of planning, design, product definition, QA, deployment, and monitoring. Simple, elegant, and scalable network design, automation, and data analytics are the keys to meeting our demands. In this role, you will be part of a team that is responsible for conceiving design solutions, developing and deploying network software, systems, and tools that keep the network operating at maximum reliability, scalability, and efficiency. This role offers an opportunity to solve scaling challenges supporting billions of people using our family of apps, as well as cutting-edge challenges in AI workloads that power new Meta products.

Responsibilities
  • Establish and implement global best practices and design new scalable network solutions
  • Conceptualize, build, and maintain automation and tools to support new product introductions, network deployment, release engineering, and operations
  • Design and develop solutions that scale across a variety of hardware platforms of network equipment
  • Lead enhancements of automation for continuous integration, validations, testing infrastructure, release, and configuration management across our global backbone, data center, and edge networks
  • Work closely with our hardware, software, and sourcing teams to develop new networking solutions and influence the future of networking and its associated infrastructure
  • Conduct thorough investigations into complex technical issues across networks, ranging from automated tooling to hardware failures and network issues
  • Develop operational process improvements and implement them in scalable, automated workflows to enhance operational efficiency
  • Help increase operational efficiency between peers and cross-functional teams by identifying roadblocks, designing and delivering automation solutions, and driving change
  • Proactively find gaps that impact multiple teams, come up with the execution plan, drive the project, and influence other teams to reach there
  • Participate in an on-call rotation to learn from real-world production challenges and take the lessons to improve current and future generation products
  • Contribute to team growth and development through peer mentorship
Minimum Qualifications
  • Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
  • 6+ years of experience planning, designing, building, and/or operating scalable and reliable systems and/or networks
  • Experience coding in at least one programming language (e.g., Python, C++, Go, Rust, or Java) and rapidly learning new development languages, software, frameworks, and APIs
  • Demonstrated knowledge of TCP, IPv4/6, Routing Protocols (one or more of BGP, MPLS, ISIS, or similar), and related network services (e.g., DHCP and DNS)
  • Experience developing and understanding network device configuration in multi-vendor environments (Juniper, Cisco, Arista, Brocade, etc.)
  • Experience in configuration and maintenance of network devices and NMS systems, or applications such as web servers, load balancers, relational databases, storage systems, and messaging systems
  • Hardware evaluation and vendor management experience
Preferred Qualifications
  • Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
  • Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
  • Knowledge of IB/RDMA/RoCE Networks, including RDMA congestion control mechanisms, AI training workloads and demands they exert on networks
  • Proven experience designing, developing, and operating distributed systems at scale, with an in-depth understanding of the challenges and opportunities in this space
  • 6+ years of experience building software solutions for managing network infrastructure, with a focus on scalability and reliability
  • Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
  • In-depth knowledge of software and network debugging, profiling, and instrumentation techniques to ensure optimal system performance
  • Experience designing and maintaining automated testing infrastructure to ensure the quality and reliability of our systems
  • Master's degree or graduate work experience in Computer Science, Computer Engineering, or a related technical field

$154,000/year to $217,000/year + bonus + equity + benefits

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Network Production Engineer, DC Frontier
Network Production Engineer, DC Frontier

Socket.dev • Menlo Park (CA)

On-site
USD 184,000 - 257,000
Network Production Engineer, Network.AI
Network Production Engineer, Network.AI

Meta • Menlo Park (CA)

Hybrid
USD 210,000 - 320,000
Network Production Engineer, Delivery Engineering
Network Production Engineer, Delivery Engineering

Meta • Menlo Park (CA)

On-site
USD 180,000 - 240,000
Network Production Engineer, DC Frontier
Network Production Engineer, DC Frontier

Meta • Menlo Park (CA)

On-site
USD 184,000 - 257,000
Equity
Bonus
Network Production Engineer, Infrastructure
Network Production Engineer, Infrastructure

Meta • Menlo Park (CA)

Hybrid
USD 122,000 - 181,000
Network Production Engineer
Network Production Engineer

Meta • Menlo Park (CA)

On-site
USD 154,000 - 217,000
Network Production Engineer — Automation & Scale
Network Production Engineer — Automation & Scale

Meta • Menlo Park (CA)

Hybrid
USD 180,000 - 240,000
Network Engineer, Deployment & Support
Network Engineer, Deployment & Support

Meta • Aiken (SC)

On-site
USD 140,000 - 210,000
Network Engineer, Engineering R&D Environments
Network Engineer, Engineering R&D Environments

Socket.dev • Town of Texas (WI)

On-site
USD 193,000 - 271,000
Bonus
Equity
Benefits
Network Engineer, Foundation & Support
Network Engineer, Foundation & Support

Meta • Denver (CO)

On-site
USD 162,000 - 227,000
Equity
Comprehensive benefits