AI Software Intern

Tenstorrent

Santa Clara (CA)

On-site

USD 69,000 - 96,000

Part time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Tenstorrent is offering an on-site internship opportunity across multiple teams, including Kernels, Models, Inference, Scale-Out, Runtime, and Compiler & Infra. You will work on high-performance kernel development, model optimization, and deployment strategies for AI workloads on Tenstorrent hardware.

The program targets students pursuing BS, MS, or PhD in CS/CE/Physics/Math with a focus on parallel processing, ML, or distributed systems.

Qualifications

  • Pursuing BS, MS, or PhD in a related field.
  • Experience with ML frameworks and model deployment.
  • Familiarity with high-performance GPU/CPU programming.

Responsibilities

  • Develop high-performance kernels for Tenstorrent hardware.
  • Optimize ML models for hardware execution and efficiency.
  • Collaborate across Kernels, Models, Inference, Scale Out, Runtime, Compiler & Infra teams.

Skills

Kernel optimization
High performance coding
Parallel processing
End-user ML intuition

Education

BS/MS/PhD in CS/CE/Physics/Math

Tools

PyTorch
CUDA
Python

Job description

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities.

Overview

This is the team that makes “it runs on Tenstorrent hardware” actually true for real models, at real scale, and builds the infrastructure that keeps the rest of engineering moving fast.

This role is on-site based out of Austin, TX or Santa Clara, CA.

What you might work on

This posting spans multiple teams within Kernels, Models, Inference, Scaleout & our Runtime teams. One application, one recruiter screen, then we match you to the specific team and location that fits best. Teams include:

  • Kernels: Develops high performance kernels on Tenstorrent hardware
  • Models: Optimizes ML models (LLMs, vision models, video and image generation, and other architectures) for our hardware
  • Inference Server: development and serving-side optimization
  • Runtime: Builds the software engine that manages memory, task scheduling, and code execution on hardware while an application is actively running.
  • Scale Out: communication and coordination between devices in a distributed AI system.
  • Compiler & Infra: Create tools that optimize AI models into high-performance programs on Tenstorrent hardware, covering memory planning, profiling, debug, and emulation.
Who You Are
  • Currently pursuing a BS, MS, or PhD in Comp Sci, Comp Eng, Physics/Math or a related field.
  • Coursework or projects any of the following: parallel processing, machine learning or distributed systems.
  • A high comfort level as an end user of AI and agentic flows.
What We Need
  • Experience with model quantization, kernel fusion, or other optimization techniques; or low level high performance programming.
  • Hands on experience working with Pytorch, experimenting with and deploying models, for inference or training.
  • Experience with C/C++ or kernel development in languages like CUDA, or Python and at least one ML framework (PyTorch, TensorFlow, JAX).
  • Knowledge of RTL, HDL is nice to have on certain teams.
What You Will Learn
  • How models get adapted to run well on specialized hardware
  • How to write high performance code that can get the most out of the underlying hardware
  • Production ML serving infrastructure at scale
  • How infrastructure choices affect the velocity of an entire engineering org
USA Hiring Timelines

This internship opportunity is available throughout our 3 terms with the following corresponding recruitment cycles:

  • Winter Term: Jan–Apr work term, Sept–Dec recruit.
  • Summer Term: May–Aug work term, Oct–Apr recruit.
  • Fall Term: Sept–Dec work term, Jan–Aug recruit.

Please note these timelines are for reference only. Actual timelines may vary.

Compensation for all interns at Tenstorrent ranges from $50/hr - $70/hr including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made.

Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer.

This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI SW Intern, Cloud, Infrastructure & Data Centre Deployment
AI SW Intern, Cloud, Infrastructure & Data Centre Deployment

Tenstorrent • Austin (CA)

On-site
USD 69,000 - 96,000
AI Software Intern
AI Software Intern

Tenstorrent University Jobs • Austin (TX), Santa Clara (CA)

On-site
USD 69,000 - 96,000
Competitive compensation
Equal opportunity employer
Software Engineer, TT-Distributed
Software Engineer, TT-Distributed

Tenstorrent • Town of Texas (WI)

On-site
USD 100,000 - 500,000
Hardware Intern - Architecture, AI HW & System on a Chip
Hardware Intern - Architecture, AI HW & System on a Chip

Tenstorrent • Santa Clara (CA)

On-site
USD 69,000 - 96,000
Applied AI Engineer
Applied AI Engineer

Tenstorrent • Santa Clara (CA)

Hybrid
USD 100,000 - 500,000
Applied AI Engineer
Applied AI Engineer

Tenstorrent • Santa Clara (CA)

On-site
USD 100,000 - 500,000
Technical Program Manager
Technical Program Manager

Tenstorrent • United States

On-site
USD 100,000 - 500,000
Technical Program Manager
Technical Program Manager

Tenstorrent • Santa Clara (CA)

Hybrid
USD 100,000 - 500,000
Director, Strategy & Solutions
Director, Strategy & Solutions

Tenstorrent • California (MO)

Hybrid
USD 100,000 - 500,000
Director, Applications & Solutions GTM
Director, Applications & Solutions GTM

Tenstorrent • United States

Hybrid
USD 100,000 - 500,000