ML Software Intern: Deploy & Optimize Models on CIM Chips

TetraMem - Accelerate The World

San Jose (CA)

On-site

USD 48,000 - 62,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TetraMem invites a Software Intern to contribute to tools, frameworks, and applications enabling neural networks on our novel analog compute-in-memory chips. This internship sits at the intersection of software, hardware, and AI.

You will develop and optimize Python or C++ code for model compression, deployment, and runtime environments, and adapt ML models to improve efficiency on compute-in-memory hardware.

Qualifications

  • Pursuing degree in CS, EE, or related field
  • Strong programming experience in Python or C++
  • Understanding of data structures, algorithms, and software architecture
  • Familiarity with AI/ML frameworks (e.g., TensorFlow, PyTorch) is a plus
  • Eagerness to learn and grow in a fast-paced environment

Responsibilities

  • Develop and optimize Python or C++ code for neural network and model deployment, compression, and runtime environments
  • Analyze and adapt ML models to improve compatibility and efficiency on compute-in-memory hardware and software
  • Support software design, development, and performance profiling
  • Collaborate with AI researchers and hardware engineers to validate system-level functionality
  • Participate in code reviews, testing, and documentation

Skills

Python
C++
Data structures
Algorithms
Software architecture
TensorFlow
PyTorch

Education

Bachelor's degree in Computer Science, Electrical Engineering, or related field

Job description

TetraMem invites a Software Intern to contribute to tools, frameworks, and applications enabling neural networks on our novel analog compute-in-memory chips. This internship sits at the intersection of software, hardware, and AI.

You will develop and optimize Python or C++ code for model compression, deployment, and runtime environments, and adapt ML models to improve efficiency on compute-in-memory hardware.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

US 2026 Software - Machine Learning Intern
US 2026 Software - Machine Learning Intern

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 48,000 - 62,000
ML Compiler Engineer for AI Hardware Accelerators
ML Compiler Engineer for AI Hardware Accelerators

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 160,000 - 300,000
Analog Hardware Intern - Compute-in-Memory Circuits
Analog Hardware Intern - Compute-in-Memory Circuits

TetraMem Inc. • San Jose (CA), Northern (KY)

Hybrid
USD 48,000 - 62,000
ML Systems Intern: Supercomputing & Hardware Co-Design
ML Systems Intern: Supercomputing & Hardware Co-Design

Etched.ai, Inc. • San Jose (CA)

On-site
12-week paid internship
Generous housing support
Daily lunch and dinner provided
ML Accelerator Performance Tools Intern
ML Accelerator Performance Tools Intern

Etched • San Jose (CA)

On-site
USD 20,000 - 31,000
Embedded ML Inference & Optimization Engineer
Embedded ML Inference & Optimization Engineer

Applied Intuition Inc. • Sunnyvale (CA)

Hybrid
USD 159,000 - 200,000
ML System Software Intern — Supercomputing & Hardware Co-Design
ML System Software Intern — Supercomputing & Hardware Co-Design

The Consensus • San Jose (CA)

On-site
Housing support
Lunch and dinner provided
Direct mentorship from industry leads
Performance Tools Intern — Profiling a Next-Gen ML Accelerator
Performance Tools Intern — Profiling a Next-Gen ML Accelerator

The Consensus • San Jose (CA)

On-site
USD 34,000 - 62,000
Intern DRAM Design Engineer — AI-Enabled Memory Systems
Intern DRAM Design Engineer — AI-Enabled Memory Systems

Micron Technology • Boise (ID)

On-site
USD 30,000 - 41,000
Medical, dental, and vision plans
Paid time-off
Paid holidays
+1
AI/CAD Engineer Intern: ML for Chip Design
AI/CAD Engineer Intern: ML for Chip Design

Falcomm • Atlanta (GA)

On-site
USD 27,552 - 41,328
Competitive Salary
Sick Leave