Get more replies from employers
Send a job-specific resume in minutes.
The Institute of Foundation Models in Sunnyvale, California, seeks an intern to help optimize large-scale AI systems on NVIDIA GPUs. You will work alongside researchers in a fast-paced environment, contributing to performance modeling and simulator development.
This internship requires strong programming skills in C++, CUDA, and knowledge of NVIDIA GPU architecture. Ideal candidates are pursuing degrees in Computer Science or related fields and have a passion for AI and performance engineering.
The Institute of Foundation Models is dedicated to advancing the science and engineering of large-scale AI systems. Our researchers and engineers develop cutting‑edge foundation models while pushing the limits of high‑performance computing and efficient AI inference. By combining deep expertise in machine learning, systems engineering, and hardware optimization, we build scalable AI solutions that drive scientific discovery and real‑world impact.
As part of the team, interns work alongside world‑class researchers and performance engineers to optimize the execution of large-scale foundation models on next‑generation NVIDIA GPU architectures. This internship provides hands‑on experience in low-level GPU performance analysis, kernel optimization, and hardware‑aware inference acceleration.
This intensive internship offers a unique opportunity to contribute to the development of a simulator and profiling framework for foundation model inference on NVIDIA GPUs.
Responsibilities include:
Currently pursuing a degree in Computer Science, Computer Engineering, Electrical Engineering, Artificial Intelligence, High-Performance Computing, or a related quantitative discipline.