Senior Network Reliability Engineer – Cloud & DC Ops Equity

NVIDIA Corporation

Santa Clara (CA)

Hybrid

USD 136,000 - 264,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA Corporation is seeking a Senior Network Reliability Engineer to support and maintain cloud and data center networks powering our software stack from graphics drivers to AI workloads. You will triage incidents, remediate alerts, and engage with vendors to resolve hardware and software issues while driving improvements in change management and operations.

The ideal candidate has 5+ years in network operations, strong TCP/IP and BGP/OSPF knowledge, CSP experience, and hands-on automation

Qualifications

  • 5+ years of experience in network operations and reliability.
  • Deep knowledge of TCP/IP, BGP, OSPF, MPLS, IS-IS, VxLAN, EVPN, QoS, GRE, IPsec, DNS, and MACsec.
  • Experience driving incident management and alert response within SLAs.
  • Experience with one or more CSP environments: AWS, Azure, GCP, OCI.
  • Hands-on with tooling and automation for provisioning, monitoring, and managing complex network infrastructures.

Responsibilities

  • Engage in 24/7 global shift rotations to provide remote support for network repairs and changes.
  • Drive operational improvements in change management and daily operations by following procedures.
  • Manage and operate large scale IP network technologies and infrastructures.
  • Utilize skills in Peering and Datacenter interconnect technologies: PNI, Transit, Exchange, Passive DWDM, Wave circuits.
  • Monitor and support the network health of on-premises and cloud infrastructures.

Skills

TCP/IP
BGP
OSPF
MPLS
EVPN
VxLAN
IPsec
DNS
MACsec
Troubleshooting
Incident management
CSP environments
Arista
Juniper
Fortinet
Python
Shell scripting
Linux
Netbox
Nautobot
Prometheus
Grafana
Panoptes

Education

Bachelor’s degree in Computer Science, related technical field, or equivalent experience

Tools

Arista
Juniper
Fortinet
Netbox
Nautobot
Prometheus
Grafana
Panoptes

Job description

NVIDIA Corporation is seeking a Senior Network Reliability Engineer to support and maintain cloud and data center networks powering our software stack from graphics drivers to AI workloads. You will triage incidents, remediate alerts, and engage with vendors to resolve hardware and software issues while driving improvements in change management and operations.

The ideal candidate has 5+ years in network operations, strong TCP/IP and BGP/OSPF knowledge, CSP experience, and hands-on automation

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SDN Systems Engineer — Cloud Networking & Equity
Senior SDN Systems Engineer — Cloud Networking & Equity

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Health insurance
401(k) plan
+1
Senior Cloud Network Architect
Senior Cloud Network Architect

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 334,000
Senior SDN Systems Engineer - Cloud Networking, Equity
Senior SDN Systems Engineer - Cloud Networking, Equity

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Equity
Senior Cloud Network Architect - Large-Scale & Automation
Senior Cloud Network Architect - Large-Scale & Automation

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 168,000 - 333,000
Equity
Benefits
Senior SDN Systems Engineer — Cloud Networking & CI/CD
Senior SDN Systems Engineer — Cloud Networking & CI/CD

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior Network Engineer - DGX Cloud
Senior Network Engineer - DGX Cloud

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 168,000 - 333,000
Equity
Benefits
Senior Cloud Operations Engineer: Automation & Reliability Lead
Senior Cloud Operations Engineer: Automation & Reliability Lead

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Comprehensive benefits package
Senior SDN Systems Engineer - Cloud Networking Architect
Senior SDN Systems Engineer - Cloud Networking Architect

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior Cloud Reliability & Automation Engineer
Senior Cloud Reliability & Automation Engineer

NVIDIA • California (MO)

On-site
USD 184,000 - 288,000
Senior Network Reliability Engineer - DGX Cloud
Senior Network Reliability Engineer - DGX Cloud

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 136,000 - 265,000