Senior Full-Stack Lead Engineer

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 224,000 - 357,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA Corporation is hiring a Senior Full-Stack Software Engineer for the DGX Cloud AI Infrastructure team. This role focuses on architecture across frontend, backend, and data layers to deliver scalable web products.

You will own multi-team initiatives, drive reliability, and mentor engineers to improve code quality, testing, and observability. Requires 12+ years of production web systems experience, cloud expertise, and strong collaboration.

Qualifications

  • Bachelor’s degree or higher in Computer Science or a related technical field (or equivalent experience).
  • Strong cross-functional collaboration skills and ability to translate complex use cases into technical requirements.
  • Experience with cloud platforms (AWS, GCP, or Azure), containers, orchestration (Docker, Kubernetes), and CI/CD.

Responsibilities

  • Lead architecture and delivery of high-scale web products across frontend, backend, and data layers.
  • Own multi-team initiatives end-to-end with RFCs, design reviews, rollouts and success metrics.
  • Improve reliability, performance, and observability to meet exascale standards.
  • Mentor engineers and promote code quality, testing, security, and observability.

Job description

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 30 years. Today, we're at the forefront of AI innovation powering breakthroughs in research, autonomous vehicles, robotics, and more. The DGX Cloud team builds and operates the AI infrastructure that fuels this progress. We’re looking for a Senior Full-Stack Software Engineer to join the AI Hub team within the DGX Cloud AI Infrastructure organization. The AI Hub team accelerates AI research by ensuring NVIDIA’s AI infrastructure is used efficiently, transparently, and at scale. Our primary goal is to build a unified, self-service “single pane of glass” portal that enables AI researchers to efficiently manage, monitor, and optimize their use of Managed AI research Superclusters.

What You’ll Be Doing:
  • Lead the architecture and delivery of high-scale web products across frontend, backend services, and data layers, with clear availability and latency targets (SLOs/SLAs).
  • Own multi-team initiatives end to end: problem discovery, RFCs/design reviews, phased rollouts, and success metrics tied to product and business outcomes.
  • Drive reliability, performance, and observability improvements to meet exascale standards.
  • Establish engineering standards and reusable platforms/design systems to reduce complexity, support load and long-term tech debt.
  • Collaborate with NVIDIA AI Research teams to identify pain points and deliver the next generation user experience that accelerates their work.
  • Mentor and sponsor engineers; improve code quality, testing, security, and observability through reviews, pairing, and coaching.
  • Stay ahead of AI/ML infrastructure trends and drive adoption of best practices within the team.
What We Need To See:
  • 12+ years of software engineering experience delivering production web systems.
  • Bachelor’s degree or higher in Computer Science or a related technical field (or equivalent experience).
  • Strong cross-functional collaboration skills, including active listening, translating complex use cases into clear technical requirements, and designing data models aligned with business logic and outcomes.
  • Deep cloud expertise (AWS, GCP, or Azure), infrastructure as code, containers, and orchestration (Docker, Kubernetes), along with mature CI/CD and safe deployment practices.
  • Full-stack depth: modern SPA frameworks (React/Next.js or Vue/Nuxt), JavaScript/TypeScript, and one or more backend languages (Node.js, Python, and/or Golang).
  • Familiarity with observability stacks such as OpenSearch, Prometheus, Grafana, or Loki.
  • Proficiency in API design (REST), schema evolution, and integration patterns, with a strong commitment to automated testing.
  • Experience building machine learning platforms or self-service internal infrastructure tools focused on efficiency, resiliency, and observability.
  • Clear written and verbal communication skills, strong problem-solving ability, and a growth mindset.
  • Experience leveraging AI-assisted development tools (e.g., Cursor).
Ways to Stand Out from the Crowd:
  • Hands‑on ML platform depth (MLE experience or strong familiarity with DL frameworks such as PyTorch, TensorFlow, JAX; distributed training ecosystems like Ray).
  • Datacenter‑scale operational experience, including GPU cluster debugging, performance triage, and root‑cause analysis across complex distributed systems.

At NVIDIA, you’ll be immersed in a diverse, supportive environment where you’re empowered to do your best work.

The DGX Cloud AI Infrastructure team is at the core of NVIDIA’s AI efforts building the software that makes scalable research possible.

Join us and help power the next wave of innovation.

NVIDIA provides competitive salaries and a comprehensive benefits package.

Our engineering teams are expanding rapidly due to exceptional growth.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

The base salary range is 224,000 USD - 356,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 13, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.

As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing.

Today, our AI infrastructure powers global intelligence, transforming every industry.

Learn more about NVIDIA.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Full-Stack Lead Engineer
Senior Full-Stack Lead Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 224,000 - 356,500
Equity
Comprehensive benefits package
Senior Customer Success Engineer - DGX Cloud
Senior Customer Success Engineer - DGX Cloud

NVIDIA Corporation • Santa Clara (CA)

Remote
USD 200,000 - 322,000
Distinguished Engineer, System Software Integration
Distinguished Engineer, System Software Integration

NVIDIA Gruppe • California (MO)

On-site
USD 320,000 - 489,000
Equity
Benefits package
Distinguished Engineer, System Software Integration
Distinguished Engineer, System Software Integration

NVIDIA • California (MO)

On-site
USD 320,000 - 489,000
Equity
Benefits package
Senior DGX Cloud AI Infrastructure Software Engineer
Senior DGX Cloud AI Infrastructure Software Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Technical Marketing Engineer - DSX AI Infrastructure Software
Senior Technical Marketing Engineer - DSX AI Infrastructure Software

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

On-site
USD 160,000 - 322,000
Equity
Benefits package
Senior Data Platform and Analytics Product Manager - DGX Cloud
Senior Data Platform and Analytics Product Manager - DGX Cloud

NVIDIA Corporation • Kansas

Hybrid
USD 208,000 - 328,000
Equity
Benefits package
Senior Software Engineer, DGX Cloud AI Infrastructure
Senior Software Engineer, DGX Cloud AI Infrastructure

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Data Platform and Analytics Product Manager - DGX Cloud
Senior Data Platform and Analytics Product Manager - DGX Cloud

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 208,000 - 328,000
Equity
Benefits
Hybrid work model
Senior Solutions Architect, Generative AI
Senior Solutions Architect, Generative AI

NVIDIA Corporation • Santa Clara (CA)

Remote
USD 184,000 - 357,000
Equity
Benefits