Senior Staff Platform Engineer

NVIDIA

Santa Clara (CA)

On-site

USD 200,000 - 322,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is seeking a Senior Staff Platform Engineer to architect, build, and scale foundational infrastructure for demanding AI/ML workloads. This hands-on technical leadership role spans cloud, networks, storage, and content delivery, with end-to-end ownership from design to production readiness and large-scale adoption.

The role requires deep expertise in Linux/Unix, distributed systems, and automation, with strong Python/Go and IaC experience.

Qualifications

  • Bachelor's degree or equivalent experience in a related technical field.
  • 12+ years of industry experience in relevant roles.
  • Proven success architecting, building, and operating large-scale distributed platforms or infrastructure systems in production.
  • Strong technical depth in cloud infrastructure, distributed systems, networking, compute, storage, platform engineering, or content delivery.
  • Deep knowledge of Linux/Unix, TCP/IP, DNS, TLS, HTTP/S, proxies, load balancing, availability, scalability, and fault-tolerant design.
  • Strong programming and automation skills with Python or Go, plus hands-on experience with infrastructure-as-code and orchestration.
  • Experience with AWS, Azure, or Google Cloud Platform and ability to troubleshoot complex systems across layers.
  • Demonstrated ability to independently drive architecture across multiple teams and mentor other engineers.

Responsibilities

  • Architect, build, and operate highly available platform services for AI/ML, distributed compute, and data-intensive workloads.
  • Own end-to-end delivery from architecture through production readiness and scale.
  • Collaborate with Cloud, Networking, Security, AI/ML, and infrastructure teams to solve cross-domain challenges.
  • Advance CDN and edge infrastructure, including caching, origin design, TLS, WAF, DNS, and traffic management.

Skills

Distributed systems
Cloud infrastructure
Networking
Python
Go
Linux/Unix
Docker

Education

Bachelor's degree in Computer Science, Computer Engineering, Electrical Engineering, or related technical field

Tools

Terraform
Kubernetes
AWS
Azure
GCP

Job description

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

Ready to build the platforms that make AI at scale possible? NVIDIA is seeking a Senior Staff Platform Engineer to architect, build, and scale foundational infrastructure for some of our most demanding compute and AI/ML workloads. The role spans distributed systems, cloud, networking, content delivery, automation, and reliability engineering. Success means turning complex infrastructure challenges into resilient platforms that can grow with NVIDIA’s rapidly evolving needs. This is a hands‑on technical leadership role with broad influence across Cloud, Networking, Security, AI/ML, and Developer Infrastructure teams.

What You Will Be Doing:

This role focuses on architecting, building, and scaling highly available platform services for AI/ML, distributed compute, and data‑intensive workloads. The work spans cloud, compute, GPU infrastructure, networking, storage, DNS, load balancing, proxies, traffic management, and content delivery, with end‑to‑end ownership from architecture and design through implementation, production readiness, and large‑scale adoption. A key part of the role is advancing CDN and edge infrastructure, including HTTP caching, origin architecture, TLS, WAF, rate limiting, traffic routing, and global load balancing. The role also drives automation, infrastructure‑as‑code, and self‑service capabilities using technologies such as Python, Go, Kubernetes, and Terraform. Observability, capacity analytics, incident learnings, and performance data are used to continuously improve reliability, scalability, efficiency, and operational simplicity. Close collaboration with Cloud, Networking, Security, AI/ML, and infrastructure teams is central to solving complex problems that span multiple technology domains.

What We Need To See:
  • Bachelor’s degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field, or equivalent experience

  • 12+ years of relevant industry experience.

  • Proven success architecting, building, and operating large‑scale distributed platforms or infrastructure systems in production.

  • Strong technical depth in several areas such as cloud infrastructure, distributed systems, networking, compute, storage, platform engineering, or content delivery.

  • Deep knowledge of Linux/Unix, TCP/IP, DNS, TLS, HTTP/S, proxies, load balancing, availability, scalability, and fault‑tolerant system design.

  • Strong programming and automation skills with Python, Go, or similar languages, plus hands‑on experience with infrastructure‑as‑code and orchestration.

  • Experience with AWS, Azure, or Google Cloud Platform and the ability to troubleshoot complex systems across application, operating system, network, and infrastructure layers.

  • Demonstrated ability to independently drive architecture and implementation across multiple teams, communicate effectively in complex situations, and mentor other engineers.

Ways To Stand Out From The Crowd:
  • Experience building platforms for AI/ML training, inference, model serving, GPU‑accelerated workloads, distributed compute, or high‑performance computing.

  • Deep expertise with CDN and edge platforms such as Akamai, AWS CloudFront, Fastly, or Cloudflare, including caching, origin design, WAF, DNS, TLS, and global traffic management.

  • Experience developing self‑service platform capabilities that enable engineering teams to consume infrastructure reliably and at scale.

  • Proven use of SLIs, SLOs, error budgets, capacity analytics, and reliability metrics to deliver measurable improvements.

  • Experience distributing models, datasets, containers, software artifacts, or other large objects across globally distributed environments.

Want to work on infrastructure where scale, AI, networking, and reliability come together? Join us!

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 200,000 USD - 322,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 23, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff Platform Engineer
Senior Staff Platform Engineer

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Senior Staff Platform Engineer
Senior Staff Platform Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Comprehensive benefits
Senior Staff Client Platform Engineer
Senior Staff Client Platform Engineer

Visa Hunt • United States

On-site
USD 200,000 - 322,000
Equity
Benefits
Senior Platform Architect, Senior Platform Architect
Senior Platform Architect, Senior Platform Architect

NVIDIA • Santa Clara (CA)

On-site
USD 168,000 - 271,000
Comprehensive benefits package
Equity eligibility
Highly competitive salaries
Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Staff Platform Engineer, Design Automation
Staff Platform Engineer, Design Automation

NVIDIA • Santa Clara (CA)

On-site
USD 196,000 - 311,000
Equity
Benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • United States

On-site
USD 168,000 - 322,000
Equity
Benefits
Solutions Architect, AI Factory Infrastructure DevOps
Solutions Architect, AI Factory Infrastructure DevOps

NVIDIA • Austin (TX)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Equity
Benefits