GPU Systems Research Intern: Optimize ML Kernels

Together

San Francisco (CA)

On-site

USD 80,000 - 96,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Housing stipend
Competitive compensation

Job summary

Together AI, the AI Native Cloud, offers an internship for a Systems Research Engineer Intern focused on GPU programming. You will co-design kernels and model architectures to boost AI system performance and efficiency, collaborating with hardware, software, and modeling teams.

Our program runs 12 to 14 weeks, with internships in May–August or June–September. We provide housing stipends and competitive compensation, reflecting location and role.

Qualifications

  • Strong background in GPU programming and parallel computing (CUDA or Triton).
  • Knowledge of ML/AI applications and models.
  • Familiar with performance profiling and optimization tools for GPUs.
  • Excellent problem-solving and analytical skills.

Responsibilities

  • Optimize and fine-tune GPU code to improve performance and scalability.
  • Collaborate with cross-functional teams to integrate GPU-accelerated solutions into software.
  • Stay updated on latest GPU programming techniques and technologies.

Skills

GPU programming
ML/AI knowledge
Parallel computing
Profiling tools
Problem solving

Tools

CUDA
Triton

Job description

Together AI, the AI Native Cloud, offers an internship for a Systems Research Engineer Intern focused on GPU programming. You will co-design kernels and model architectures to boost AI system performance and efficiency, collaborating with hardware, software, and modeling teams.

Our program runs 12 to 14 weeks, with internships in May–August or June–September. We provide housing stipends and competitive compensation, reflecting location and role.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

GPU Systems Research Intern – High-Perf ML Kernels
GPU Systems Research Intern – High-Perf ML Kernels

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipend
GPU Systems Research Intern: High-Performance AI
GPU Systems Research Intern: High-Performance AI

Together • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
GPU Systems Research Intern: Accelerate AI Performance
GPU Systems Research Intern: Accelerate AI Performance

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
GPU Systems Research Intern (CUDA/Triton)
GPU Systems Research Intern (CUDA/Triton)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
GPU Systems Research Intern — CUDA/Triton, On‑Site SF
GPU Systems Research Intern — CUDA/Triton, On‑Site SF

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipend
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
AI Model Optimization Engineer Intern (GPU & Frameworks)
AI Model Optimization Engineer Intern (GPU & Frameworks)

Advanced Micro Devices • Austin (TX), Northern (KY)

Hybrid
USD 28,000 - 41,000