GPU Systems Research Intern: High-Performance AI

Together

San Francisco (CA)

On-site

USD 80,000 - 96,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Housing stipend
Competitive benefits

Job summary

Together AI in San Francisco is seeking a Systems Research Engineer Intern focused on GPU programming to develop and optimize GPU-accelerated kernels for ML/AI applications. You will co-design kernels with the modeling team and work with hardware/software groups to improve AI system performance.

The internship runs on-site from January to April at the SF HQ, offering exposure to cutting-edge GPU research and collaborative engineer environments.

Qualifications

  • Strong background in GPU programming and parallel computing.
  • Familiarity with CUDA and/or Triton architectures and APIs.
  • Knowledge of ML/AI applications and models.
  • Experience with performance profiling and GPU optimization tools.

Responsibilities

  • Optimize and fine-tune GPU kernels for ML/AI workloads.
  • Collaborate with cross-functional teams to integrate GPU-accelerated solutions.
  • Stay current with new GPU programming techniques and hardware trends.

Skills

GPU programming
Parallel computing
CUDA / Triton
ML/AI familiarity

Job description

Together AI in San Francisco is seeking a Systems Research Engineer Intern focused on GPU programming to develop and optimize GPU-accelerated kernels for ML/AI applications. You will co-design kernels with the modeling team and work with hardware/software groups to improve AI system performance.

The internship runs on-site from January to April at the SF HQ, offering exposure to cutting-edge GPU research and collaborative engineer environments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Systems Research Intern: Accelerate AI Performance
GPU Systems Research Intern: Accelerate AI Performance

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
GPU Systems Research Intern – High-Perf ML Kernels
GPU Systems Research Intern – High-Perf ML Kernels

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipend
GPU Systems Research Intern — CUDA/Triton, On‑Site SF
GPU Systems Research Intern — CUDA/Triton, On‑Site SF

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
GPU Systems Research Intern: Optimize ML Kernels
GPU Systems Research Intern: Optimize ML Kernels

Together • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
GPU Systems Research Intern (CUDA/Triton)
GPU Systems Research Intern (CUDA/Triton)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipend
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
AI Systems & GPU Performance Engineer Intern
AI Systems & GPU Performance Engineer Intern

Advanced Micro Devices • San Jose (CA), Santa Clara (CA)

Hybrid
USD 22,000 - 45,000