Senior Technical Support Engineer - Ethernet and AI Infrastructure

NVIDIA

United States

Remote

USD 150,000 - 190,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

NVIDIA is seeking a Senior Technical Support Engineer to own complex investigations in Ethernet networking and AI infrastructure. You will be a trusted technical advisor to strategic customers and collaborate with Engineering, Product, Marketing, and Support to resolve issues and improve practices.

You will manage high-severity incidents, drive root-cause analysis, and guide cross-functional teams to resolution, while enhancing automation and documentation for scalable support.

Qualifications

  • 5+ years of experience providing in-depth technical support for networking, hardware, software, or enterprise infrastructure.
  • Deep knowledge of Ethernet and data center networking, including TCP/IP, Layer 2/3 technologies, ARP, STP, LACP, MLAG, IGMP, PIM, BGP, OSPF, routing and switching.
  • Proven experience troubleshooting complex network issues using tcpdump, Wireshark, packet generators, telemetry, and log-analysis tools.
  • Strong Linux system administration and networking skills, including diagnosing server, OS, driver, hardware, and performance issues.
  • Excellent customer-facing, written and verbal communication, with ability to explain complex topics and manage expectations.
  • Proven ability to lead high-severity customer situations and coordinate multiple technical teams.
  • Experience using agentic AI technologies to improve troubleshooting and automation.
  • Hands-on AI infrastructure experience with data centers, distributed systems, virtualization, or HPC.

Responsibilities

  • Investigate and resolve complex issues involving NVIDIA Ethernet networking, Linux servers, and large-scale AI infrastructure.
  • Own critical customer issues from initial response through resolution, coordinating cross-functional teams.
  • Reproduce problems, analyze diagnostics, identify root causes, and resolve installation, operation, maintenance, performance issues.
  • Serve as trusted advisor to enterprise, cloud, and service-provider customers via phone, email, and conference support.
  • Collaborate with Engineering/Product to translate field findings into product improvements and docs.
  • Provide technical leadership during high-severity incidents, managing priorities and customer expectations.

Skills

Ethernet networking
Linux systems
Troubleshooting
Customer communication
Technical leadership
Cross-functional collaboration
Agentic AI tools

Education

Bachelor's degree in CS/CE/EE or related

Tools

tcpdump
Wireshark
Telemetry tools
NVIDIA Spectrum-X

Job description

We are seeking a highly motivated Senior Technical Support Engineer with deep expertise in Ethernet networking and AI infrastructure. In this highly visible, customer-facing role, you will own complex technical investigations and critical escalations involving large-scale data center and AI environments.

As a member of our NVEX Global Technical Support team, you will serve as a trusted technical advisor to strategic customers. The ideal candidate combines strong hands‑on troubleshooting skills with excellent customer communication, crisis management, and technical leadership. You will collaborate closely with Engineering, Product Management, Marketing, and Support teams to resolve issues and improve our products and support practices.


What you'll be doing
  • Investigate and resolve complex issues involving NVIDIA Ethernet networking, Linux systems, servers, and large-scale AI infrastructure, including end-to-end solutions such as NVIDIA Spectrum-X.
  • Own critical customer issues from initial response through resolution, coordinating cross-functional teams and providing clear, timely communication to customers and leadership.
  • Reproduce customer problems, analyze diagnostic data, identify root causes, and resolve issues involving installation, operation, maintenance, performance, and multivendor interoperability.
  • Serve as a trusted technical advisor to enterprise, cloud, and service‑provider customers through telephone, email, and conference‑based support engagements.
  • Partner with Engineering and Product teams to translate field findings into product improvements, support tools, technical documentation, and troubleshooting methodologies.
  • Provide technical leadership during high‑severity incidents, remaining calm and decisive while managing priorities, risks, and customer expectations.
  • Improve support effectiveness through automation, knowledge sharing, structured debugging practices, and the practical use of agentic AI tools.

What we need to see
  • 5+ years of experience providing in-depth technical support and debugging for networking, hardware, software, or enterprise infrastructure products.
  • Deep knowledge of Ethernet and data center networking, including TCP/IP, Layer 2 and Layer 3 technologies, ARP, STP, LACP, MLAG, IGMP, PIM, BGP, OSPF, routing, and switching.
  • Proven experience troubleshooting complex network issues using tcpdump, Wireshark, packet generators, telemetry, and log‑analysis tools.
  • Strong Linux system administration and networking skills, including the ability to diagnose server, operating‑system, driver, hardware, and performance issues.
  • Excellent customer‑facing, written, verbal, and presentation skills, with the ability to explain complex technical issues, manage expectations, and build trust with customers and executives.
  • Proven ability to lead high‑severity customer situations, coordinate multiple technical teams, make sound decisions under pressure, and drive issues to timely resolution.
  • Strong analytical, organizational, and prioritization skills, with the ability to work independently and manage multiple complex issues.
  • Practical experience using agentic AI technologies such as Claude, Codex, or Cursor to improve troubleshooting, automation, documentation, or day‑to‑day productivity.
  • Hands‑on experience with AI infrastructure and at least two of the following areas: data centers, distributed systems, accelerated servers, virtualization, deep learning frameworks, Docker, Kubernetes, or high‑performance computing - advantage.
  • A bachelor's or master's degree in Computer Science, Computer Engineering, Electrical Engineering, Networking, or a related discipline, or equivalent practical experience.

Ways to stand out from the crowd
  • Experience troubleshooting large‑scale AI clusters, high‑performance computing environments, cloud infrastructure, or hyperscale data centers.
  • Expertise in advanced networking technologies such as VXLAN, EVPN, RoCE, RDMA, congestion control, quality of service, and lossless Ethernet.
  • Knowledge of AI and HPC technologies such as GPUs, NCCL, MPI, Slurm, distributed training frameworks, and workload orchestration.
  • Experience with NVIDIA networking, accelerated computing, Spectrum‑X, or NVIDIA AI Enterprise technologies.
  • Proficiency with Python, Bash, Ansible, YAML, APIs, or similar technologies used for diagnostics and automation.

We are looking for a technically accomplished support engineer who can solve difficult infrastructure problems, earn customer trust, and lead effectively during critical situations. If you are passionate about Ethernet networking, AI infrastructure, and delivering an outstanding customer experience, we would like to hear from you.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA • Nashville (TN)

On-site
USD 108,000 - 172,500
Equity options
Comprehensive benefits package
Inclusive work environment
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA Gruppe • Town of Texas (WI)

On-site
USD 120,000 - 207,000
Equity options
Comprehensive benefits
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA Corporation • Northern (KY)

Hybrid
USD 108,000 - 173,000
Equity
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA • Redmond (WA)

On-site
USD 108,000 - 173,000
Equity
Benefits
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA Corporation • Tennessee

Hybrid
USD 108,000 - 207,000
Equity
Benefits
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA Corporation • Town of Texas (WI)

Hybrid
USD 120,000 - 207,000
Equity
Benefits
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA Gruppe • Tennessee

On-site
USD 108,000 - 207,000
Equity options
Comprehensive benefits package
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA • Austin (TX)

On-site
USD 120,000 - 207,000
Equity
Comprehensive benefits
Senior HPC Support Engineer - Ethernet and AI Infrastructure
Senior HPC Support Engineer - Ethernet and AI Infrastructure

NVIDIA Corporation • Redmond (WA)

On-site
USD 108,000 - 207,000
Equity
Benefits
Senior Ethernet & AI Infrastructure Support Engineer
Senior Ethernet & AI Infrastructure Support Engineer

NVIDIA • United States

Remote
USD 150,000 - 190,000