Senior Cloud AI Platform Engineer - GPU & Serverless

NVIDIA Corporation

Santa Clara, Northern (CA, KY)

Hybrid

USD 184,000 - 288,000

Full time

9 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA Corporation in Santa Clara, California, is seeking an AI/ML Engineer for its Cloud Functions team. This role focuses on a serverless deployment platform enabling AI applications and scalable inference across GPU-backed, cloud-agnostic Kubernetes clusters.

You will drive expert-level AI model delivery, contribute to platform features, and mentor teams while ensuring performance and reliability at scale.

Qualifications

  • Masters, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Applied Math, or related field.
  • At least 2 years work experience with Python, Rust, Golang, Linux or Bash.
  • Experience in Deep Learning and Machine Learning; expertise in using AI/DL frameworks and inferencing software such as SGLang, vLLM, TensorRT-LLM, or Dynamo.
  • Knowledge of CPU and GPU architecture.
  • Excellent interpersonal skills including ability to explain sophisticated technical topics to non-experts.
  • Experience in the design, implementation, and release of AI/ML products to market.
  • A flexible technologist familiar with all aspects of the software development lifecycle.

Responsibilities

  • Becoming a trusted subject matter expert by understanding user challenges and constraints.
  • Translate this into product requirements and solutions, accelerating delivery of AI models and inference hosted on the NVCF platform.
  • Leading implementation of key features.
  • Conducting user-acceptance testing, load testing and performance evaluations.
  • Emphasis on customer experience, performance optimization and platform reliability.
  • Mentoring and embedding with other engineering teams building products on top of our platform on best practices for AI/ML workloads at scale with excellent performance, including ML reliability engineering at scale.
  • Producing reference architectures to guide customer use cases, applying the newest NVCF product features with the latest AI technologies.
  • Shepherding customer issues to resolution and providing timely warning of issues and risks.
  • Evaluating new and innovative technologies and tooling as the AI-at-scale landscape evolves to ensure we have a competitive product and forward-looking roadmap.

Skills

Python
Rust
Golang
Linux
Bash
Deep Learning
Machine Learning
CPU/GPU architecture
Communication

Education

Master's / PhD or equivalent

Tools

SGLang
vLLM
TensorRT-LLM
Dynamo

Job description

NVIDIA Corporation in Santa Clara, California, is seeking an AI/ML Engineer for its Cloud Functions team. This role focuses on a serverless deployment platform enabling AI applications and scalable inference across GPU-backed, cloud-agnostic Kubernetes clusters.

You will drive expert-level AI model delivery, contribute to platform features, and mentor teams while ensuring performance and reliability at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Cloud Platform Engineer - Serverless & GPU
Senior AI Cloud Platform Engineer - Serverless & GPU

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior Cloud Platform Engineer AI-Driven, Scalable Backend
Senior Cloud Platform Engineer AI-Driven, Scalable Backend

NVIDIA Corporation • Northern (KY)

Hybrid
USD 168,000 - 270,000
Senior Cloud AI Infra Engineer: Go/C, Kubernetes, GPUs
Senior Cloud AI Infra Engineer: Go/C, Kubernetes, GPUs

Nvidia • Santa Clara (CA)

On-site
USD 240,000 - 340,000
Competitive salaries
Comprehensive benefits package
Equity
Serverless AI Platform Lead Engineer — Equity
Serverless AI Platform Lead Engineer — Equity

NVIDIA AI • Seattle (WA)

On-site
USD 140,000 - 190,000
Equity
Comprehensive benefits package
Senior Cloud Engineer - AI Developer Tools & CUDA SaaS
Senior Cloud Engineer - AI Developer Tools & CUDA SaaS

NVIDIA AI • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Cloud Systems Engineer — Scientific Computing Platform
Senior Cloud Systems Engineer — Scientific Computing Platform

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Platform Engineer - AI/ML Infra & Global CDN
Senior Platform Engineer - AI/ML Infra & Global CDN

Nvidia Corporation • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Comprehensive benefits
Remote Senior Cloud Data Platform Engineer
Remote Senior Cloud Data Platform Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior AI Cloud Solutions Architect—GPU Platform
Senior AI Cloud Solutions Architect—GPU Platform

NVIDIA AI • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior Cloud AI Tools Engineer — CUDA, SaaS, Equity
Senior Cloud AI Tools Engineer — CUDA, SaaS, Equity

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits