Systems Software Engineer - AI and Cloud

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 124,000 - 242,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA Corporation in Santa Clara invites a System Software Engineer - AI and Cloud to join a collaborative team in Silicon Valley. You will work on cloud-native AI deployments, microservices, and advanced models using NVIDIA frameworks.

You will design RAG workflows, evaluate AI solutions, and produce tutorials and demos for industry conferences, while collaborating with product, hardware, and software teams.

Qualifications

  • Bachelor’s or Master’s in Software Engineering, Computer Science, Computer Engineering, Electrical Engineering or related field, or equivalent experience.
  • 3+ years of experience in software development.
  • Proficiency in Python and JavaScript with solid data structures and algorithms knowledge.
  • Basic familiarity with C++ in high-performance computing contexts.
  • Experience building cloud-native systems for Kubernetes and inference frameworks like Triton or vLLM.
  • Strong understanding of API design for scalable, production-grade inference systems.

Responsibilities

  • Evaluate cloud-native, full-stack apps using microservices for AI use cases with NVIDIA frameworks and tools.
  • Design and implement agentic workflows using techniques like Retrieval-Augmented Generation (RAG).
  • Evaluate AI solutions and compile findings into clear reports for engineering leadership.
  • Suggest product improvements to senior executives and engineering management.
  • Collaborate with product, marketing, hardware, software engineering, and QA teams to improve offerings.
  • Create developer content, tutorials, code samples, and demos for NVIDIA tools and libraries.
  • Produce technical whitepapers and product briefs and present demos at industry events.

Skills

Python
JavaScript
C++
Kubernetes
API design
LLMs

Education

Bachelor’s or Master’s in Software Engineering/CS/CE/EE

Tools

Triton Inference Server
vLLM
NVIDIA frameworks

Job description

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It is a unique legacy of innovation that is fueled by great technology—and amazing people. Today, we are tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what is never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you will be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. Join NVIDIA, where we are pushing the boundaries of what is possible in AI and cloud computing.

As a versatile System Software Engineer - AI and Cloud, you will be part of a team of dedicated professionals that thrives on innovation and collaboration. Located in the heart of Silicon Valley, you will have the opportunity to work on groundbreaking projects that craft the future of technology. This role offers an outstanding chance to engage with advanced AI models and cloud-native architectures, making significant contributions to NVIDIA's versatile products and technologies.

What you’ll be doing:
  • Evaluate cloud-native, full-stack applications using microservices architecture to power AI use cases, bringing to bear NVIDIA frameworks, SDKs, and microservices.
  • Design and implement agentic workflows with advanced techniques like Retrieval-Augmented Generation (RAG) and the latest AI models.
  • Evaluate user experiences and analyze the technical performance of AI solutions, compiling findings into comprehensive reports.
  • Offer practical suggestions for product improvement to senior executives and engineering management.
  • Engage with various teams across NVIDIA such as product, marketing, hardware, software engineering, and QA to improve NVIDIA's product offerings.
  • Develop developer-focused content, including detailed tutorials and code samples, to demonstrate the latest features in NVIDIA’s tools and libraries.
  • Write technical whitepapers and product briefs, and run technical demos of our products at prominent industry conferences.
What we need to see:
  • A Bachelor’s or Master’s in Software Engineering, Computer Science, Computer Engineering, Electrical Engineering or a related degree (or equivalent experience) 3+ years of experience.
  • Proficiency in Python and JavaScript for programming and debugging, with a strong foundation in data structures, algorithms, and software design principles.
  • Basic familiarity with C++ programming and its application in high-performance computing environments.
  • Experience in crafting cloud-native systems optimized for Kubernetes deployment, using inference frameworks such as vLLM and NVIDIA Triton Inference Server.
  • A solid understanding of API design principles for building scalable, production-ready inference systems.
Ways to stand out from the crowd:
  • Advanced knowledge of LLMs, modern AI software architecture, and cloud APIs.
  • Contributions to public-facing technical content and open-source projects.
  • Expertise in deploying LLM inference frameworks like Triton Inference Server, vLLM, or TensorRT, including on Kubernetes or edge devices to improve performance.

The base salary range is 124,000 USD - 195,500 USD for Level 2, and 152,000 USD - 241,500 USD for Level 3. You will also be eligible for equity and benefits.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Learn more about NVIDIA.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff Software Engineer — AI Applications and Platform Foundations
Senior Staff Software Engineer — AI Applications and Platform Foundations

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Inclusive work environment
Benefits package
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Durham (NC)

On-site
USD 224,000 - 357,000
Equity
Benefits
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Austin (TX)

On-site
USD 224,000 - 432,000
Equity
Benefits
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Westford (MA)

On-site
USD 224,000 - 432,000
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA Corporation • Austin (TX)

On-site
USD 152,000 - 288,000
Applied AI Engineer
Applied AI Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits package
Hybrid work model
Technical Marketing Engineer - AI Platform Software
Technical Marketing Engineer - AI Platform Software

NVIDIA • Santa Clara (CA)

On-site
USD 136,000 - 253,000
Equity
Benefits
Senior Staff Software Engineer - AI Applications and Platform Foundations
Senior Staff Software Engineer - AI Applications and Platform Foundations

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Distinguished Engineer, System Software Integration
Distinguished Engineer, System Software Integration

Nvidia Corporation • Santa Clara (CA)

On-site
USD 320,000 - 489,000
Equity
Benefits package
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Equity
Benefits