GPU Systems Research Intern (CUDA/Triton)

Togetherai

San Francisco (CA)

On-site

USD 80,000 - 96,000

Part time

5 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Housing stipend
Competitive compensation

Job summary

Together AI is seeking a Systems Research Engineer Intern specialized in GPU Programming to develop and optimize GPU-accelerated kernels for ML/AI applications. You will co-design GPU kernels and model architectures, collaborating with hardware and software teams to advance efficient GPU programming models.

You will stay current with the latest GPU programming techniques and contribute to AI infrastructure innovations, aligning with industry-leading engineering practices.

Qualifications

  • Strong background in GPU programming and parallel computing, such as CUDA and/or Triton.
  • Knowledge of ML/AI applications and models.
  • Knowledge of performance profiling and optimization tools for GPU programming.
  • Excellent problem-solving and analytical skills.

Responsibilities

  • Optimize and fine-tune GPU code to achieve better performance and scalability.
  • Collaborate with cross-functional teams to integrate GPU-accelerated solutions into existing software systems.
  • Stay up-to-date with the latest advancements in GPU programming techniques and technologies.

Skills

CUDA
Triton
ML/AI models
Profiling tools
Problem-solving

Job description

Together AI is seeking a Systems Research Engineer Intern specialized in GPU Programming to develop and optimize GPU-accelerated kernels for ML/AI applications. You will co-design GPU kernels and model architectures, collaborating with hardware and software teams to advance efficient GPU programming models.

You will stay current with the latest GPU programming techniques and contribute to AI infrastructure innovations, aligning with industry-leading engineering practices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Systems Research Intern: High-Performance AI
GPU Systems Research Intern: High-Performance AI

Together • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
GPU Systems Research Intern – High-Perf ML Kernels
GPU Systems Research Intern – High-Perf ML Kernels

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipend
GPU Systems Research Intern — CUDA/Triton, On‑Site SF
GPU Systems Research Intern — CUDA/Triton, On‑Site SF

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
GPU Systems Research Intern: Optimize ML Kernels
GPU Systems Research Intern: Optimize ML Kernels

Together • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
Systems Research Engineer Intern - GPU Programming (Summer 2027)
Systems Research Engineer Intern - GPU Programming (Summer 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipend
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Togetherai • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive benefits
Systems Research Engineer Intern - GPU Programming (Winter 2027)
Systems Research Engineer Intern - GPU Programming (Winter 2027)

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
GPU Systems Research Intern: Accelerate AI Performance
GPU Systems Research Intern: Accelerate AI Performance

Together AI • San Francisco (CA)

On-site
USD 80,000 - 87,000
Housing stipends
Remote GPU Kernel Engineer: CUDA/Triton Optimization
Remote GPU Kernel Engineer: CUDA/Triton Optimization

anyone-ai • United States

On-site
USD 138,000 - 248,000