Senior Systems Software Engineer - NV Cloud Functions

Nvidia Corporation in

Santa Clara (CA)

On-site

USD 184,000 - 288,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

NVIDIA is seeking a Senior Systems Software Engineer for the NV Cloud Functions (Finance) team to build and operate a serverless deployment platform for AI applications at scale. You will translate user challenges into requirements, lead feature implementations, and mentor other teams working on AI/ML workloads across GPU-enabled Kubernetes clusters.

The role emphasizes performance, reliability, and reference architectures, with equity and comprehensive benefits.

Qualifications

  • Masters, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Applied Math, or related field.
  • At least 2 years work experience with Python, Rust, Golang, Linux or Bash.
  • Experience in Deep Learning and Machine Learning; expertise in using AI/DL frameworks and inferencing software such as SGLang, vLLM, TensorRT-LLM, or Dynamo.
  • Knowledge of CPU and GPU architecture.
  • Excellent interpersonal skills including ability to explain sophisticated technical topics to non-experts.
  • Experience in the design, implementation, and release of AI/ML products to market. A flexible technologist familiar with all aspects of the software development lifecycle.

Responsibilities

  • Becoming a trusted subject matter expert by understanding user challenges and constraints. Translate this into product requirements and solutions, accelerating delivery of AI models and inference hosted on the NVCF platform.
  • Leading implementation of key features. Conducting user-acceptance testing, load testing and performance evaluations. Emphasis on customer experience, performance optimization and platform reliability.
  • Mentoring and embedding with other engineering teams building products on top of our platform on best practices for AI/ML workloads at scale with excellent performance, including ML reliability engineering at scale.
  • Producing reference architectures to guide customer use cases, applying the newest NVCF product features with the latest AI technologies.
  • Shepherding customer issues to resolution and providing timely warning of issues and risks.
  • Evaluating new and innovative technologies and tooling as the AI-at-scale landscape evolves to ensure we have a competitive product and forward-looking roadmap.

Skills

Python
Rust
Golang
Linux/Bash
Deep Learning
ML frameworks
CPU/GPU architecture
Communication skills
SDLC

Education

Masters/PhD or equivalent experience

Job description

Senior Systems Software Engineer - NV Cloud Functions (Finance)

NVIDIA Cloud Functions team is looking for a motivated, product-minded AI/ML Engineer with domain expertise in AI platform engineering at scale. Our team builds and operates a serverless deployment platform for enabling AI applications. Our product enables and scales AI inferencing workloads using globally distributed orchestration of workloads on GPU-backed cloud-agnostic Kubernetes clusters. You will be working with a team of passionate and skilled engineers that are continuously innovating at the speed of light to provide the best product possible, for both external customers and internal NVIDIA teams. We are looking for someone to join us at the forefront of defining cloud engineering paradigms for AI at scale.

What You'll be Doing:
  • Becoming a trusted subject matter expert by understanding user challenges and constraints. Translate this into product requirements and solutions, accelerating delivery of AI models and inference hosted on the NVCF platform.
  • Leading implementation of key features. Conducting user-acceptance testing, load testing and performance evaluations. Emphasis on customer experience, performance optimization and platform reliability.
  • Mentoring and embedding with other engineering teams building products on top of our platform on best practices for AI/ML workloads at scale with excellent performance, including ML reliability engineering at scale.
  • Producing reference architectures to guide customer use cases, applying the newest NVCF product features with the latest AI technologies.
  • Shepherding customer issues to resolution and providing timely warning of issues and risks.
  • Evaluating new and innovative technologies and tooling as the AI-at-scale landscape evolves to ensure we have a competitive product and forward-looking roadmap.
What We Need to See:
  • Masters, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Applied Math, or related field
  • At least 2 years work experience with Python, Rust, Golang, Linux or Bash.
  • Experience in Deep Learning and Machine Learning; expertise in using AI/DL frameworks and inferencing software such as SGLang, vLLM, TensorRT-LLM, or Dynamo.
  • Knowledge of CPU and GPU architecture.
  • Excellent interpersonal skills including ability to explain sophisticated technical topics to non-experts.
  • Experience in the design, implementation, and release of AI/ML products to market. A flexible technologist familiar with all aspects of the software development lifecycle.
Ways to stand out from the crowd:
  • Demonstrate a strong desire to share knowledge with clients, partners and co-workers, able to show this through previous work.
  • Demonstrate expertise through projects or Open Source contributions in HPC, Data Analytics, Machine Learning, Deep Learning, Cloud Native Projects, Kubernetes, Slurm, or enabling GPU workloads.
  • Show a willingness and ability to dig into unfamiliar territories to tackle complex problems through examples in previous work.
  • Prior experience in building distributed systems.

NVIDIA offers highly competitive salaries and a comprehensive benefits package. NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most brilliant and talented people in the world working for us. Are you a creative engineer with a drive for advancing the state of AI and bringing it to the cloud? If you love to tackle problems and advocate for continuous, innovative improvement, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Systems Software Engineer - NV Cloud Functions
Senior Systems Software Engineer - NV Cloud Functions

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 184,000 - 288,000
Equity
Benefits
Senior Solutions Architect, NVIDIA Cloud Partners
Senior Solutions Architect, NVIDIA Cloud Partners

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Cloud Software Engineer, DGXC Data Services
Senior Cloud Software Engineer, DGXC Data Services

Visa Hunt • United States

On-site
USD 184,000 - 288,000
Principal Solutions Architect, NVIDIA Cloud Partners
Principal Solutions Architect, NVIDIA Cloud Partners

Visa Hunt • United States

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior Solutions Architect, NVIDIA Cloud Partners - Telco
Senior Solutions Architect, NVIDIA Cloud Partners - Telco

NVIDIA • Seattle (WA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Inclusive work environment
Senior System Software Engineer - Scientific Computing PaaS
Senior System Software Engineer - Scientific Computing PaaS

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Manager, Software Engineering - DGX Cloud
Manager, Software Engineering - DGX Cloud

Nvidia Corporation in • Austin (TX)

On-site
USD 272,000 - 431,000
Senior Cloud Software Engineer, Developer Tools
Senior Cloud Software Engineer, Developer Tools

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior System Software Engineer - Scientific Computing PaaS
Senior System Software Engineer - Scientific Computing PaaS

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Solutions Architect, NVIDIA Cloud Partners
Senior Solutions Architect, NVIDIA Cloud Partners

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits