Senior Systems Software Engineer - NV Cloud Functions

NVIDIA Corporation

Santa Clara, Northern (CA, KY)

Hybrid

USD 184,000 - 288,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA Corporation in Santa Clara, California, is seeking an AI/ML Engineer for its Cloud Functions team. This role focuses on a serverless deployment platform enabling AI applications and scalable inference across GPU-backed, cloud-agnostic Kubernetes clusters.

You will drive expert-level AI model delivery, contribute to platform features, and mentor teams while ensuring performance and reliability at scale.

Qualifications

  • Masters, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Applied Math, or related field.
  • At least 2 years work experience with Python, Rust, Golang, Linux or Bash.
  • Experience in Deep Learning and Machine Learning; expertise in using AI/DL frameworks and inferencing software such as SGLang, vLLM, TensorRT-LLM, or Dynamo.
  • Knowledge of CPU and GPU architecture.
  • Excellent interpersonal skills including ability to explain sophisticated technical topics to non-experts.
  • Experience in the design, implementation, and release of AI/ML products to market.
  • A flexible technologist familiar with all aspects of the software development lifecycle.

Responsibilities

  • Becoming a trusted subject matter expert by understanding user challenges and constraints.
  • Translate this into product requirements and solutions, accelerating delivery of AI models and inference hosted on the NVCF platform.
  • Leading implementation of key features.
  • Conducting user-acceptance testing, load testing and performance evaluations.
  • Emphasis on customer experience, performance optimization and platform reliability.
  • Mentoring and embedding with other engineering teams building products on top of our platform on best practices for AI/ML workloads at scale with excellent performance, including ML reliability engineering at scale.
  • Producing reference architectures to guide customer use cases, applying the newest NVCF product features with the latest AI technologies.
  • Shepherding customer issues to resolution and providing timely warning of issues and risks.
  • Evaluating new and innovative technologies and tooling as the AI-at-scale landscape evolves to ensure we have a competitive product and forward-looking roadmap.

Skills

Python
Rust
Golang
Linux
Bash
Deep Learning
Machine Learning
CPU/GPU architecture
Communication

Education

Master's / PhD or equivalent

Tools

SGLang
vLLM
TensorRT-LLM
Dynamo

Job description

NVIDIA Cloud Functions team is looking for a motivated, product-minded AI/ML Engineer with domain expertise in AI platform engineering at scale. Our team builds and operates a serverless deployment platform for enabling AI applications. Our product enables and scales AI inferencing workloads using globally distributed orchestration of workloads on GPU-backed cloud-agnostic Kubernetes clusters. You will be working with a team of passionate and skilled engineers that are continuously innovating at the speed of light to provide the best product possible, for both external customers and internal NVIDIA teams. We are looking for someone to join us at the forefront of defining cloud engineering paradigms for AI at scale.

Are you a creative engineer with a drive for advancing the state of AI and bringing it to the cloud?

What You'll be Doing:
  • Becoming a trusted subject matter expert by understanding user challenges and constraints.
  • Translate this into product requirements and solutions, accelerating delivery of AI models and inference hosted on the NVCF platform.
  • Leading implementation of key features.
  • Conducting user-acceptance testing, load testing and performance evaluations.
  • Emphasis on customer experience, performance optimization and platform reliability.
  • Mentoring and embedding with other engineering teams building products on top of our platform on best practices for AI/ML workloads at scale with excellent performance, including ML reliability engineering at scale.
  • Producing reference architectures to guide customer use cases, applying the newest NVCF product features with the latest AI technologies.
  • Shepherding customer issues to resolution and providing timely warning of issues and risks.
  • Evaluating new and innovative technologies and tooling as the AI-at-scale landscape evolves to ensure we have a competitive product and forward-looking roadmap.
What We Need to See:
  • Masters, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Applied Math, or related field.
  • At least 2 years work experience with Python, Rust, Golang, Linux or Bash.
  • Experience in Deep Learning and Machine Learning; expertise in using AI/DL frameworks and inferencing software such as SGLang, vLLM, TensorRT-LLM, or Dynamo.
  • Knowledge of CPU and GPU architecture.
  • Excellent interpersonal skills including ability to explain sophisticated technical topics to non-experts.
  • Experience in the design, implementation, and release of AI/ML products to market.
  • A flexible technologist familiar with all aspects of the software development lifecycle.
Ways to stand out from the crowd:
  • Demonstrate a strong desire to share knowledge with clients, partners and co-workers, able to show this through previous work.
  • Demonstrate expertise through projects or Open Source contributions in HPC, Data Analytics, Machine Learning, Deep Learning, Cloud Native Projects, Kubernetes, Slurm, or enabling GPU workloads.
  • Show a willingness and ability to dig into unfamiliar territories to tackle complex problems through examples in previous work.
  • Prior experience in building distributed systems.
Benefits and Compensation:
  • NVIDIA offers highly competitive salaries and a comprehensive benefits package.
  • You will also be eligible for equity and benefits.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

Applications for this job will be accepted at least until September 13, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.

As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing.

Today, our AI infrastructure powers global intelligence, transforming every industry.

Learn more about NVIDIA.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Software Engineer, DGXC Data Services
Senior Cloud Software Engineer, DGXC Data Services

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity
Benefits
Distinguished Engineer, Scaled Out Inferencing
Distinguished Engineer, Scaled Out Inferencing

Nvidia Corporation • Santa Clara (CA)

On-site
USD 320,000 - 489,000
Equity
Benefits package
Distinguished Engineer, System Software Integration
Distinguished Engineer, System Software Integration

NVIDIA • California (MO)

On-site
USD 320,000 - 489,000
Equity
Benefits package
NCX Senior Engineer
NCX Senior Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits
Systems Software Engineer - AI and Cloud
Systems Software Engineer - AI and Cloud

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 124,000 - 242,000
Equity
Benefits
Senior Staff Platform Engineer
Senior Staff Platform Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Comprehensive benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • Colorado

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Platform AI Engineer
Senior Platform AI Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Distinguished Engineer, Scaled Out Inferencing
Distinguished Engineer, Scaled Out Inferencing

NVIDIA • California (MO)

On-site
USD 320,000 - 489,000
Equity
Benefits
Solutions Architect - NVIDIA Cloud Partners
Solutions Architect - NVIDIA Cloud Partners

NVIDIA • United States

On-site
USD 184,000 - 357,000
Equity
Benefits