Stand out for this role — generate a tailored resume and cover letter in about a minute.
As a Network Reliability Engineer on the OCI Network Availability team, you will play a crucial role in ensuring the high availability and performance of Oracle Cloud's global network infrastructure. This role involves applying engineering methodologies to measure, monitor, and automate the reliability of OCI’s network, supporting millions of users across a vast, distributed environment.
You will be part of a fast-paced, innovative team responsible for swiftly responding to network disruptions, identifying root causes, and collaborating with both internal and external stakeholders to restore services. Your work will also focus on automating daily operations, improving workflow efficiency, and optimizing network performance. With OCI's expansive global footprint, you will manage hundreds of thousands of network devices across a mix of dedicated backbone infrastructure, CLoS networks, and the internet.
Experience: Experience working in a large-scale ISP or cloud provider environment, supporting global network infrastructure is a plus. Prior experience in a network operations role, with a proven track record of handling complex network events.
Technical Skills: Strong proficiency in network protocols and services, including MPLS, BGP, OSPF, IS-IS, TCP/IP, IPv4/IPv6, DNS, DHCP, VxLAN, and EVPN. Experience with network automation, scripting, and data center design. Python is preferred, though expertise in other scripting or compiled languages is a plus. Hands-on experience with network monitoring and telemetry solutions, with the ability to leverage these tools to drive improvements in network reliability. Familiarity with network modeling and programming, including YANG, OpenConfig, and NETCONF.
Problem-Solving and Collaboration: Ability to apply engineering principles to resolve complex network issues, collaborating across teams to deliver effective solutions.
Communication: Strong communication skills, both written and verbal, with the ability to present technical information clearly to both technical and non-technical stakeholders. Demonstrated experience in influencing product roadmap decisions, priorities, and feature development through sound judgment and technical expertise.
This role involves participation in an on-call rotation, providing 24/7 support for critical network events and incidents. You will have the opportunity to work in a highly dynamic environment with exposure to cutting-edge technologies and large-scale cloud infrastructure.
Referrals increase your chances of interviewing at Oracle by 2x