Senior Software Engineer, Cosmos Infrastructure and End to End Performance

NVIDIA AI

Santa Clara (CA)

On-site

USD 180,000 - 260,000

Full time

9 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA AI seeks a highly skilled professional to build state-of-the-art World Foundation Models and drive end-to-end performance analysis for data center and edge deployments. You will develop infrastructure to automate data ingestion, curation, pre-training, and deployment for Physical AI systems.

The role requires a Master’s degree in a STEM field and 5+ years of experience with large-scale parallel distributed systems and AI workload optimization, with proficiency in Python, C++, and

Qualifications

  • Requires a Master's degree in a STEM field and 5 years of experience with large-scale parallel distributed systems and AI workload optimization.
  • Proficiency in Python, C++, and Distributed PyTorch is essential, along with a deep understanding of World Foundation Models.

Responsibilities

  • Build state-of-the-art World Foundation Models and drive end-to-end performance analysis for data center and edge deployments.
  • Develop infrastructure to automate data ingestion, curation, pre-training, and deployment for Physical AI systems.

Skills

Distributed PyTorch
Python
C/C++
CUDA
Performance Modeling
Computer Architecture
Distributed Systems
World Foundation Models
Quantization
Data Curation
Transformer Models
Cloud Service Providers
DNNs
Networking
Storage Systems
Edge Deployment

Education

Master's degree in STEM

Job description

Build state-of-the-art World Foundation Models and drive end-to-end performance analysis for data center and edge deployments. Develop infrastructure to automate data ingestion, curation, pre-training, and deployment for Physical AI systems.

Requirements

Requires a Master's degree in a STEM field and 5 years of experience with large-scale parallel distributed systems and AI workload optimization. Proficiency in Python, C++, and Distributed PyTorch is essential, along with a deep understanding of World Foundation Models.

Key Skills
  • Distributed PyTorch
  • Python
  • C/C++
  • CUDA
  • Performance Modeling
  • Computer Architecture
  • Distributed Systems
  • World Foundation Models
  • Quantization
  • Data Curation
  • Transformer Models
  • Cloud Service Providers
  • DNNs
  • Networking
  • Storage Systems
  • Edge Deployment
Benefits
  • Equity, Benefits
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cosmos Infra Engineer - End-to-End Performance
Senior Cosmos Infra Engineer - End-to-End Performance

NVIDIA AI • Santa Clara (CA)

On-site
USD 180,000 - 260,000
Equity
Benefits
Senior Software Engineer, Cosmos Infrastructure and End to End Performance
Senior Software Engineer, Cosmos Infrastructure and End to End Performance

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Software Engineer, Cosmos Infrastructure and End to End Performance
Senior Software Engineer, Cosmos Infrastructure and End to End Performance

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Software Engineer, Cosmos Infrastructure and End to End Performance
Senior Software Engineer, Cosmos Infrastructure and End to End Performance

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Cosmos Infrastructure & E2E Performance Engineer
Senior Cosmos Infrastructure & E2E Performance Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Cosmos Infra & End-to-End Performance Engineer — Equity
Cosmos Infra & End-to-End Performance Engineer — Equity

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Cosmos Infra & End-to-End Performance Engineer
Senior Cosmos Infra & End-to-End Performance Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Deep Learning Engineer
Senior Deep Learning Engineer

NVIDIA • Indiana (PA)

On-site
USD 100,000 - 150,000
Applied Scientist -- Foundation Models, SSO
Applied Scientist -- Foundation Models, SSO

Amazon Web Services (AWS) • Santa Clara (CA)

On-site
USD 172,000 - 222,000
Member of Technical Staff (Performance Optimization)
Member of Technical Staff (Performance Optimization)

Fireworks AI • United States

On-site
USD 150,000 - 260,000