GPU Systems Research Intern – High-Perf ML Kernels

Together AI

San Francisco (CA)

On-site

USD 80,000 - 87,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Housing stipend

Job summary

Together AI, the AI Native Cloud, is seeking a Systems Research Engineer Intern specialized in GPU programming. You will develop and optimize GPU-accelerated kernels and algorithms for ML/AI applications, co-design GPU kernels and model architecture, and collaborate with hardware and software teams to advance efficient GPU programming models.

The internship runs 12 to 14 weeks with dates May 17–Aug 6 or Jun 14–Sept 3.

Qualifications

  • Strong background in GPU programming and parallel computing (CUDA/Triton).
  • Knowledge of ML/AI models and applications.
  • Experience with performance profiling/optimization tools for GPUs.
  • Excellent problem-solving and analytical skills.

Responsibilities

  • Optimize and fine-tune GPU code to improve performance and scalability.
  • Collaborate with cross-functional teams to integrate GPU-accelerated solutions into existing software systems.
  • Stay up-to-date with the latest GPU programming techniques and technologies.

Skills

GPU programming
Parallel computing
ML/AI knowledge

Tools

CUDA
Triton
Profiling tools

Job description

Together AI, the AI Native Cloud, is seeking a Systems Research Engineer Intern specialized in GPU programming. You will develop and optimize GPU-accelerated kernels and algorithms for ML/AI applications, co-design GPU kernels and model architecture, and collaborate with hardware and software teams to advance efficient GPU programming models.

The internship runs 12 to 14 weeks with dates May 17–Aug 6 or Jun 14–Sept 3.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Systems Research Intern: Optimize ML Kernels
GPU Systems Research Intern: Optimize ML Kernels

Together • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
GPU Systems Research Intern: High-Performance AI
GPU Systems Research Intern: High-Performance AI

Together • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
GPU Systems Research Intern (CUDA/Triton)
GPU Systems Research Intern (CUDA/Triton)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
GPU Systems Research Intern: Accelerate AI Performance
GPU Systems Research Intern: Accelerate AI Performance

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
GPU Systems Research Intern — CUDA/Triton, On‑Site SF
GPU Systems Research Intern — CUDA/Triton, On‑Site SF

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipend
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
PhD AI Systems & GPU Performance Intern - Co-op (Hybrid)
PhD AI Systems & GPU Performance Intern - Co-op (Hybrid)

AMD • Santa Clara (CA)

Hybrid
USD 55,000 - 96,000