Senior Software Engineer, Network Visibility Platform - DGX Cloud

NVIDIA

Santa Clara (CA)

On-site

USD 168,000 - 322,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

NVIDIA in Santa Clara, CA seeks a Senior Software Engineer for the Global Network Visibility team. Lead a platform turning topology, telemetry, and measurement into trusted, self-service insight to accelerate incident triage and service availability.

You will define durable interfaces, drive cross-domain decisions, and mentor engineers while balancing performance, cost, and reliability. Expect cutting-edge tech in a dynamic environment.

Qualifications

  • Bachelor’s degree or equivalent experience.
  • 10+ years of architecting and operating large-scale software or infrastructure platforms.
  • Strong proficiency in Python, Go, or a comparable systems language.
  • Deep expertise across network topology, telemetry, observability, or related areas.
  • Hands-on experience with Prometheus, Grafana, OpenTelemetry and Kubernetes.

Responsibilities

  • Set the architecture for a unified platform spanning topology, configuration, telemetry, active measurement, change intelligence, and self-service experiences.
  • Build an authoritative network model connecting physical topology, logical overlays, configurations, and service dependencies.
  • Define a telemetry & measurement strategy with reachability tests, streaming signals, traffic data, and change events.
  • Deliver health, path, blast-radius, and self-service exploration products for customer self-service.
  • Establish interfaces, data contracts, quality standards, and operating mechanisms for cross-team contributions.
  • Drive cross-functional roadmaps, balance speed, cost, reliability, and mentor engineers.

Skills

Python
Go
Leadership
Product & operational judgment
Cross-functional collaboration
Network topology

Education

Bachelor's degree

Tools

Prometheus
Grafana
OpenTelemetry
Kubernetes

Job description

As a Senior Software Engineer on NVIDIA’s Global Network Visibility (GNV) team within NVIDIA's Global Network Infrastructure (GNI) organization, you will lead the platform that turns network topology, configuration, telemetry, and direct measurement into trusted, self-service insight. Your work will help teams understand network health and customer impact, accelerate incident triage, and bring AI, storage, backbone, and edge infrastructure online with confidence.

You will set direction across authoritative network data, scalable measurement, and customer-facing visibility products. You will define durable interfaces, lead cross-domain decisions, and guide build-versus-buy choices. Success means the platform is credible, adopted, and trusted.

What you'll be doing:
  • Set the architecture for a unified platform spanning topology, configuration, telemetry, active measurement, change intelligence, and self-service experiences
  • Build an authoritative network model connecting physical topology, logical overlays, configurations, and service dependencies across systems of record, and use it to catch physical and logical configuration errors during cluster bring-up
  • Define a telemetry & measurement strategy combining reachability tests, streaming signals, traffic & capacity data, and change & maintenance events
  • Deliver intuitive health, path, blast-radius, and self-service exploration products that help customers answer network questions independently
  • Establish interfaces, data contracts, quality standards, and operating mechanisms that let teams contribute without fragmenting the experience
  • Drive cross-functional roadmaps; make clear tradeoffs among speed, cost, reliability, and maintainability; and mentor engineers through critical designs
What we need to see:
  • Bachelor’s degree or equivalent experience, plus 10+ years of relevant industry experience
  • Record of architecting and operating large-scale software or infrastructure platforms, with strong proficiency in Python, Go, or a comparable systems language
  • Deep expertise in one area—with breadth across the others: network topology and sources of truth, telemetry and active measurement, or observability products
  • Strong data center and backbone networking fundamentals across physical connectivity, routing, overlays, services, and customer-visible failures
  • Hands‑on experience across an observability stack such as Prometheus, Grafana, and OpenTelemetry, including active or synthetic monitoring and containerized deployment on Kubernetes
  • Experience with large-scale graph, relational, time-series, streaming, API, or event-driven systems that reconcile multiple sources
  • Technical leadership across teams, with strong product and operational judgment and clear tradeoffs among performance, resiliency, usability, and cost
Ways to stand out from the crowd:
  • Led a network visibility or infrastructure intelligence platform spanning thousands of devices or hosts
  • Experience with AI or HPC networks, including RDMA, RoCE over Spectrum-X, or InfiniBand
  • Experience with network sources of truth such as NetBox or Nautobot, streaming telemetry (gRPC/gNMI), graph databases, or front‑end development in TypeScript and React
  • SRE or network operations experience, plus a record of setting standards, making build-versus-buy decisions, or contributing to open source

NVIDIA is widely considered one of the world's most desirable employers in technology. We have some of the world's most forward-thinking and passionate people working for us. If you're creative and autonomous, we want to hear from you!

NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for over 25 years. It’s a unique legacy of innovation fueled by great technology—and dynamic people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 270,250 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 27, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Network Visibility Platform - DGX Cloud
Senior Software Engineer, Network Visibility Platform - DGX Cloud

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Network Engineer - DGX Cloud
Senior Network Engineer - DGX Cloud

NVIDIA • Santa Clara (CA)

On-site
USD 168,000 - 265,000
Senior Software Engineer, Networking DGX Cloud
Senior Software Engineer, Networking DGX Cloud

NVIDIA Gruppe • United States

On-site
USD 200,000 - 391,000
Equity and benefits
Senior Software Engineer - Traffic and Networking
Senior Software Engineer - Traffic and Networking

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity
Benefits
Senior Network Engineer - DGX Cloud
Senior Network Engineer - DGX Cloud

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 334,000
Senior Software Engineer - DGX Cloud Services and Software
Senior Software Engineer - DGX Cloud Services and Software

Socket.dev • Santa Clara (UT)

On-site
USD 168,000 - 270,000
Senior Network Security Architect
Senior Network Security Architect

NVIDIA • Santa Clara (CA)

On-site
USD 196,000 - 380,000
Equity
Comprehensive benefits
Principal Software Engineer - Networking - DGX Cloud
Principal Software Engineer - Networking - DGX Cloud

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Senior Network Security Architect
Senior Network Security Architect

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 196,000 - 380,000
Equity
Benefits
Senior Software Engineer - Traffic and Networking
Senior Software Engineer - Traffic and Networking

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 431,000