GPU Inference Performance Intern & Modeling

Institute of Foundation Models

Sunnyvale (CA)

On-site

USD 34,440 - 55,104

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hands-on experience with advanced AI systems

Job summary

The Institute of Foundation Models in Sunnyvale, California, seeks an intern to help optimize large-scale AI systems on NVIDIA GPUs. You will work alongside researchers in a fast-paced environment, contributing to performance modeling and simulator development.

This internship requires strong programming skills in C++, CUDA, and knowledge of NVIDIA GPU architecture. Ideal candidates are pursuing degrees in Computer Science or related fields and have a passion for AI and performance engineering.

Qualifications

  • Currently pursuing a degree in a relevant field.
  • Strong programming skills in C++, CUDA, and Python are required.
  • Understanding of NVIDIA GPU architecture and memory hierarchy.

Responsibilities

  • Develop analytical performance models for GPU kernels.
  • Build and validate a simulator for NVIDIA GPUs.
  • Document findings and provide actionable recommendations.

Skills

CUDA programming
GPU kernel development
Performance profiling tools
Python
C++

Education

Degree in Computer Science or related field

Tools

NVIDIA Nsight Systems
NVIDIA Nsight Compute

Job description

The Institute of Foundation Models in Sunnyvale, California, seeks an intern to help optimize large-scale AI systems on NVIDIA GPUs. You will work alongside researchers in a fast-paced environment, contributing to performance modeling and simulator development.

This internship requires strong programming skills in C++, CUDA, and knowledge of NVIDIA GPU architecture. Ideal candidates are pursuing degrees in Computer Science or related fields and have a passion for AI and performance engineering.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Inference Optimization Intern: GPU Performance Modeling
Inference Optimization Intern: GPU Performance Modeling

Ifm Us • Sunnyvale (CA)

On-site
USD 30,000 - 60,000
Inference Optimization Intern – Performance Modeling
Inference Optimization Intern – Performance Modeling

Institute of Foundation Models • Sunnyvale (CA)

On-site
Hands-on experience with advanced AI systems
Inference Optimization Intern – Performance Modeling
Inference Optimization Intern – Performance Modeling

Ifm Us • Sunnyvale (CA)

On-site
USD 30,000 - 60,000
GPU Inference Performance Engineer — Equity & Optimization
GPU Inference Performance Engineer — Equity & Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Inference Performance Engineer: AI GPU Optimization&Equity
Inference Performance Engineer: AI GPU Optimization&Equity

NVIDIA • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Equity
Benefits package
Senior AI Inference Performance Engineer — Scale GPUs
Senior AI Inference Performance Engineer — Scale GPUs

NVIDIA • California (MO)

On-site
USD 124,000 - 196,000
Equity eligibility
Senior Inference Performance Engineer — Equity & Hybrid
Senior Inference Performance Engineer — Equity & Hybrid

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Senior Inference Performance Engineer - GPU & CUDA
Senior Inference Performance Engineer - GPU & CUDA

inference.net • San Francisco (CA)

Hybrid
USD 220,000 - 320,000
AI Infrastructure Intern — Inference at Scale
AI Infrastructure Intern — Inference at Scale

DeepInfra • Palo Alto (CA)

On-site
Remote GPU Performance Engineer: Scale Training & Inference
Remote GPU Performance Engineer: Scale Training & Inference

Reka • United States

Remote
USD 120,000 - 150,000
Five weeks of paid leave
Comprehensive healthcare benefits
Visa support for H1B and OPT transfers