Stand out for this role — generate a tailored resume and cover letter in about a minute.
Ll Oefentherapie in Nashville, TN is seeking a Senior Network Operations Engineer to monitor, troubleshoot, and optimize a global network infrastructure. You will analyze data, triage incidents, and coordinate with vendors to maintain high availability.
You’ll lead automation efforts, mentor junior engineers, and participate in on-call rotations to ensure reliable 24/7 operations for Oracle customers and services worldwide.
Senior Network Operations Engineer
Location: Nashville, TN
NOTE: This position is not eligible for sponsorship.
In this role, you will monitor and troubleshoot network events, collect and analyze technical data, triage and mitigate incidents, coordinate escalations, and help drive continuous operational improvement. You will work alongside experienced engineers, partner teams, and vendors to maintain and optimize the infrastructure that supports Oracle customers and services worldwide.
Our mission is to keep OCI’s global network highly available, performant, and resilient—while delivering exceptional service to customers and dependable operational support to our engineering and technical teams.
For GNOC engineers, that mission translates into a broad, fast-moving role with real operational impact. You will help centrally manage OCI’s network infrastructure, respond to and resolve complex events, and develop automated solutions that reduce recurring operational work and improve reliability at scale.
Use established procedures and operational tooling to plan, implement, and safely complete network changes.
Mentor, onboard, and train junior network engineers.
Participate in operational rotations and provide break/fix and incident-response support.
Identify and triage actionable incidents through monitoring systems; analyze and mitigate network events; conduct or support root-cause analysis (RCA); and coordinate follow-up actions with internal support teams and vendors.
Provide on-call support as required, exercising sound independent judgment in a varied and complex operational environment.
Participate in major incident calls and use technical and analytical skills to resolve network issues affecting Oracle customers and services.
Manage fault detection, response, and escalation for OCI systems and networks, collaborating with third-party suppliers through resolution.
Collaborate with GNOC Shift Leads and management to ensure the efficient and timely completion of daily GNOC responsibilities.
Lead, contribute to, and participate in the identification, development, and evaluation of projects and tools that improve GNOC effectiveness.
Drive runbook audits and updates to maintain compliance and align operational processes with partner service teams.
Conduct interviews and participate in hiring junior-level engineers.
Lead and/or represent the GNOC in vendor meetings, service reviews, and governance boards.
Collaborate with network automation teams to integrate and improve operational support tooling.
Develop scripts and automation to reduce manual effort and improve the reliability of routine operational tasks.
Preferred experience with Python, Puppet, SQL, Ansible, network automation, and databases.
Lead technical initiatives, including the development and improvement of runbooks, methods of procedure (MOPs), operational processes, and team onboarding materials.
Support the implementation of short-, medium-, and long-term plans to achieve project objectives.
Regularly engage senior management and network leadership to ensure team priorities and project objectives are met.
Strong knowledge of networking protocols and technologies, including BGP, OSPF, IS-IS, TCP/IP, IPv4/IPv6, DNS, DHCP, MPLS, VPNs, and TLS.
Broad hands-on experience with at least three of the following: Juniper, Cisco, Arista, InfiniBand, NVIDIA, firewalls, routers, switches, circuit management, and optical/network transport services.
Strong analytical skills, including the ability to gather, correlate, and interpret data from multiple sources.
Ability to diagnose, prioritize, resolve, or appropriately escalated network alerts and faults.
Experience in a large ISP, cloud provider, or similarly complex enterprise network environment.
Exposure to commodity Ethernet hardware and networking ASICs, including Broadcom and NVIDIA/Mellanox.
Cisco, Arista and Juniper certifications are desirable.
Hands-on experience supporting carrier circuits and transport services in a 24×7 NOC, ISP, cloud-provider, data center, telecommunications, or large enterprise environment.
Demonstrated experience troubleshooting fiber, optical, Ethernet, and WAN circuit failures across multiple carriers and vendors.
Experience coordinating carrier escalations, field dispatches, remote hands, circuit turn-ups, maintenance windows, and service restoration.
Working knowledge of optical power levels, transceivers, fiber paths, cross-connects, demarcation points, interface counters, and circuit-testing methods.
Experience with DWDM, dark fiber, DIA, MPLS, VPLS, microwave, SONET, PON, or comparable transport technologies.
Ability to correlate physical-layer and circuit conditions with routing adjacencies, packet loss, latency, congestion, and customer impact.
Strong incident-management, technical documentation, vendor-management, and root-cause-analysis skills.
Experience supporting GPU and RDMA network environments is highly desirable.
Experience supporting high-performance computing (HPC) environments is highly desirable.
Experience with InfiniBand and NVIDIA networking technologies, including Spectrum, is highly desirable.
Participate in network lifecycle management, including network build, refresh, and upgrade projects.
Participate in network solution design and design-review activities.
Self-motivated, proactive, and able to work independently.
Bachelor’s degree preferred, with at least 3–5 years of relevant network operations or engineering experience.
Strong organizational, time-management, verbal, and written communication skills.
Comfortable managing a broad range of priorities in a fast-paced operational environment.
Experience with incident-response plans, processes, and strategies.
Experience supporting large-scale enterprise infrastructure and cloud computing environments in a 24/7 network operations setting, including willingness to work rotational shifts.