Inference Optimization Intern: GPU Performance Modeling

Ifm Us

Sunnyvale (CA)

On-site

USD 30,000 - 60,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Ifm Us is seeking an intern to work on the optimization of large-scale foundation models using NVIDIA GPU architectures. Interns will have the chance to collaborate with world-class researchers and engineers, gaining hands-on experience in GPU performance analysis and kernel optimization.

The ideal candidate should be pursuing a degree in Computer Science or related fields, with strong programming skills in CUDA, C++, and Python. This is a unique opportunity to contribute to cutting-edge developments in AI systems.

Qualifications

  • Currently pursuing a degree in a relevant quantitative discipline.
  • Experience with CUDA programming and GPU kernel development.
  • Understanding of NVIDIA GPU architecture and memory hierarchy.

Responsibilities

  • Develop analytical performance models for GPU kernels.
  • Build and validate a simulator for performance estimation.
  • Identify performance bottlenecks in compute and memory.

Skills

CUDA programming
GPU kernel development
C++
Python
Nsight Systems
Nsight Compute
Deep learning frameworks (PyTorch, TensorFlow)
Performance analysis

Education

Pursuing degree in Computer Science or related field

Job description

Ifm Us is seeking an intern to work on the optimization of large-scale foundation models using NVIDIA GPU architectures. Interns will have the chance to collaborate with world-class researchers and engineers, gaining hands-on experience in GPU performance analysis and kernel optimization.

The ideal candidate should be pursuing a degree in Computer Science or related fields, with strong programming skills in CUDA, C++, and Python. This is a unique opportunity to contribute to cutting-edge developments in AI systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Inference Performance Intern & Modeling
GPU Inference Performance Intern & Modeling

Institute of Foundation Models • Sunnyvale (CA)

On-site
Hands-on experience with advanced AI systems
Inference Optimization Intern – Performance Modeling
Inference Optimization Intern – Performance Modeling

Ifm Us • Sunnyvale (CA)

On-site
USD 30,000 - 60,000
Inference Optimization Intern – Performance Modeling
Inference Optimization Intern – Performance Modeling

Institute of Foundation Models • Sunnyvale (CA)

On-site
Hands-on experience with advanced AI systems
Senior LLM Inference: GPU Kernel Optimization
Senior LLM Inference: GPU Kernel Optimization

NVIDIA • Austin (TX)

On-site
USD 184,000
Equity
Benefits
GPU Systems Research Intern — AI Kernel Optimizer
GPU Systems Research Intern — AI Kernel Optimizer

Together AI • San Francisco (CA)

On-site
Competitive compensation
Housing stipends
Other benefits
GPU Inference Performance Engineer — Equity & Optimization
GPU Inference Performance Engineer — Equity & Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Senior Inference Engineer: GPU Kernel Optimizations + Equity
Senior Inference Engineer: GPU Kernel Optimizations + Equity

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Comprehensive benefits
Inference Performance Engineer: AI GPU Optimization&Equity
Inference Performance Engineer: AI GPU Optimization&Equity

NVIDIA • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Equity
Benefits package
GPU Performance Engineer | Experienced Hire
GPU Performance Engineer | Experienced Hire

Susquehanna International Group • Bala Cynwyd (PA)

On-site
USD 100,000 - 130,000
GPU Performance Engineer | Experienced Hire
GPU Performance Engineer | Experienced Hire

SIG Susquehanna • Bala Cynwyd (PA)

On-site
USD 120,000 - 160,000