Senior Storage Software Engineer, DGXC Data Services

NVIDIA

California (MO)

On-site

USD 184,000 - 287,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We solve problems in AI data management for exabyte-scale, high-performance GPU workloads.

We are seeking engineers to develop storage technologies, APIs, and observability to improve performance and reliability for AI training and inference at scale. Strong foundations in distributed systems, OS, and languages like Go, Python, Rust, C/C++, or Java

Qualifications

  • BS in Computer Science, Information Systems, Computer Engineering, or equivalent experience.
  • 5+ years of software engineering experience.
  • Strong foundation in algorithms, data structures, distributed systems, operating systems, and practical software design.
  • Experience building performance-sensitive systems, storage, backend, or cloud-native software in languages such as Go, Python, Rust, C/C++, or Java.
  • Experience with storage systems, object stores, caching, Linux systems, Kubernetes, or cloud infrastructure.
  • Ability to reason about performance, scalability, concurrency, reliability, and operational tradeoffs in production systems.
  • Ability to design APIs, document systems, communicate clearly, and break ambiguous infrastructure problems into practical execution plans.
  • Curiosity and practical judgment around AI-assisted or agentic engineering workflows, including using clear intent, specifications, acceptance criteria, tests, and verification to guide development.

Responsibilities

  • Build storage technologies, client libraries, and filesystem frameworks for AI workloads across object stores, file systems, and hybrid cloud.
  • Develop high-performance storage paths for training and inference, including data loading, checkpointing, caching, POSIX access, and object-store integration.
  • Build observability systems that diagnose storage bottlenecks and expose telemetry through production monitoring stacks.
  • Improve performance, scalability, and reliability of storage systems handling massive datasets and high-concurrency workloads.
  • Collaborate with internal AI teams, platform teams, SRE, and operations to validate storage against real workloads and production environments.
  • Use modern software engineering practices, including AI-assisted workflows, with high standards for design, testing, security, performance and verification.

Skills

Algorithms
Distributed systems
Operating systems
Go
Python
Rust
C/C++
Java
Linux
Kubernetes
Cloud infrastructure
API design
Documentation
AI-assisted workflows

Education

BS in Computer Science
Equivalent experience

Tools

Go tooling
Python tooling

Job description

The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload.

What You Will Be Doing
  • Build storage technologies, client libraries, and filesystem frameworks that help AI workloads access data across object stores, file systems, and hybrid cloud infrastructure.
  • Develop high-performance storage paths for training and inference workflows, including data loading, checkpointing, caching, POSIX-style access, and object-store integration.
  • Build observability systems that diagnose storage bottlenecks, attribute GPU idle time to I/O behavior, and expose actionable telemetry through production monitoring stacks.
  • Improve performance, scalability, and reliability of storage systems serving massive datasets, deep directory trees, and high-concurrency AI workloads.
  • Work closely with internal AI teams, platform teams, SRE, and operations to validate storage behavior against real workloads and production environments.
  • Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, performance, and verification.
What We Need To See
  • BS in Computer Science, Information Systems, Computer Engineering, or equivalent experience, with 5+ years of software engineering experience.
  • Strong foundation in algorithms, data structures, distributed systems, operating systems, and practical software design.
  • Experience building performance-sensitive systems, storage, backend, or cloud-native software in languages such as Go, Python, Rust, C/C++, or Java.
  • Experience with storage systems, object stores, caching, Linux systems, Kubernetes, or cloud infrastructure.
  • Ability to reason about performance, scalability, concurrency, reliability, and operational tradeoffs in production systems.
  • Ability to design APIs, document systems, communicate clearly, and break ambiguous infrastructure problems into practical execution plans.
  • Curiosity and practical judgment around AI-assisted or agentic engineering workflows, including using clear intent, specifications, acceptance criteria, tests, and verification to guide development.
Ways To Stand Out From The Crowd
  • Background with Linux kernel observability, eBPF, tracing, or low-overhead telemetry systems.
  • Experience with FUSE, POSIX filesystems, object-store-backed filesystems, or filesystem metadata/indexing.
  • Experience optimizing storage performance for AI training, checkpointing, inference, or large-scale data pipelines.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 10, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Storage Software Engineer, DGXC Data Services
Senior Storage Software Engineer, DGXC Data Services

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Storage Software Engineer, DGXC Data Services
Senior Storage Software Engineer, DGXC Data Services

NVIDIA • United States

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Storage Software Engineer, DGXC Data Services
Senior Storage Software Engineer, DGXC Data Services

NVIDIA • North Carolina

On-site
USD 152,000 - 242,000
Equity compensation
Benefits
Senior Storage Software Engineer, DGXC Data Services
Senior Storage Software Engineer, DGXC Data Services

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity compensation
Benefits
Senior Storage Software Engineer - DGX Cloud
Senior Storage Software Engineer - DGX Cloud

Nvidia Corporation • Santa Clara (CA)

On-site
USD 224,000 - 432,000
Equity
Benefits
Senior Storage Software Engineer - DGX Cloud
Senior Storage Software Engineer - DGX Cloud

2100 NVIDIA USA • Santa Clara (CA)

On-site
USD 224,000 - 432,000
Equity
Benefits
Senior Cloud Software Engineer, DGXC Data Services
Senior Cloud Software Engineer, DGXC Data Services

2100 NVIDIA USA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits package
Senior Engineering Manager, Object Storage - DGX Cloud
Senior Engineering Manager, Object Storage - DGX Cloud

NVIDIA AI • Santa Clara (CA)

On-site
USD 272,000 - 489,000
Senior Site Reliability Engineer - Storage
Senior Site Reliability Engineer - Storage

NVIDIA AI • Santa Clara (CA)

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior HPC Storage Engineer
Senior HPC Storage Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Comprehensive benefits package
Equity options