AI Systems & Infrastructure Intern (LLM & GPU Workloads)

Micron Technology

Austin (TX)

On-site

USD 34,000 - 44,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Micron Technology is seeking an AI Systems Software Engineering Intern to work with senior engineers on advanced systems software for LLMs and Agentic AI applications. You will profile and optimize AI inference/training workloads across GPU platforms and heterogeneous memory, interconnect, and storage systems.

The role emphasizes memory management techniques, benchmarking, and collaboration to develop AI workloads, publish results, and contribute to future platform designs.

Qualifications

  • Currently pursuing a Master’s or Ph.D. in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • Demonstrated experience with AI systems, machine learning systems, computer systems research, or systems software development through coursework, research, or projects.
  • Understanding of Large Language Models (LLMs), including transformer execution, attention mechanisms, KV cache behavior, batching, token-level latency, throughput, and memory performance considerations.
  • Proficiency in Python and C/C++, with hands-on experience developing, debugging, and optimizing software in Linux environments.
  • Experience using GPU-based performance analysis tools and at least one modern AI framework or serving stack, such as PyTorch, vLLM, TensorRT-LLM, NVIDIA Dynamo, or related technologies.

Responsibilities

  • Develop and enhance systems software, profiling tools, and experimentation frameworks for LLM training, LLM inference, and Agentic AI workloads.
  • Design, implement, and evaluate memory- and state-management techniques, including caching, tiering, compression, eviction, and lifecycle management for AI serving environments.
  • Characterize and optimize AI workload execution across GPUs, CPUs, memory subsystems, storage, and distributed infrastructure, with a focus on latency, throughput, scalability, and resource utilization.
  • Build benchmarking, simulation, and automation capabilities to evaluate data placement, migration, scheduling, and performance behavior across heterogeneous memory systems.
  • Collaborate with engineering and research teams to develop representative AI workloads, analyze experimental results, and contribute to technical publications, intellectual property, and future platform designs.

Skills

Python
C/C++
AI systems
LLMs
Linux
GPU performance tools

Education

Master's or Ph.D. in Computer Science / Computer Engineering / Electrical Engineering

Tools

PyTorch
TensorRT-LLM
vLLM
NVIDIA Dynamo

Job description

Micron Technology is seeking an AI Systems Software Engineering Intern to work with senior engineers on advanced systems software for LLMs and Agentic AI applications. You will profile and optimize AI inference/training workloads across GPU platforms and heterogeneous memory, interconnect, and storage systems.

The role emphasizes memory management techniques, benchmarking, and collaboration to develop AI workloads, publish results, and contribute to future platform designs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Systems & Infrastructure Intern for LLMs & GPUs
AI Systems & Infrastructure Intern for LLMs & GPUs

Micron • Austin (TX)

On-site
USD 34,440,000 - 48,216,000
Medical coverage
Paid time off
Paid holidays
+1
AI Infrastructure Systems Intern — Optimize LLM Workloads
AI Infrastructure Systems Intern — Optimize LLM Workloads

Micron Technology, Inc • Austin (TX)

On-site
USD 34,000 - 48,000
Medical, dental, and vision plans
Paid time off
Paid holidays
LLM AI Systems & Infrastructure Intern
LLM AI Systems & Infrastructure Intern

1000 Micron Technology, Inc. • Austin (TX)

On-site
USD 41,000 - 55,000
Medical, dental, vision plans
Paid time off
Paid holidays
LLM Inference Systems Performance Engineer
LLM Inference Systems Performance Engineer

3M HEALTHCARE • Austin (TX)

On-site
USD 150,000 - 210,000
Medical, dental, and vision coverage
Income protection benefits
Paid family leave
+1
Intern - AI Systems and Infrastructure Engineering
Intern - AI Systems and Infrastructure Engineering

Micron Technology, Inc • Austin (TX)

On-site
USD 34,000 - 48,000
Medical, dental, and vision plans
Paid time off
Paid holidays
Intern, Memory & System Architecture for AI Datacenters
Intern, Memory & System Architecture for AI Datacenters

Micron Technology, Inc • Northern (KY)

Hybrid
USD 69,000 - 77,000
Medical, dental, vision plans
Paid time off
Paid holidays
+1
AI Agentic Systems Engineer Intern — Cloud Platform
AI Agentic Systems Engineer Intern — Cloud Platform

Micron Technology, Inc • California (MO)

On-site
USD 69,000 - 77,000
Intern - AI Systems and Infrastructure Engineering
Intern - AI Systems and Infrastructure Engineering

Micron • Austin (TX)

On-site
USD 34,440,000 - 48,216,000
Medical coverage
Paid time off
Paid holidays
+1
Intern - AI Systems and Infrastructure Engineering
Intern - AI Systems and Infrastructure Engineering

1000 Micron Technology, Inc. • Austin (TX)

On-site
USD 41,000 - 55,000
Medical, dental, vision plans
Paid time off
Paid holidays
Intern - AI Systems and Infrastructure Engineering
Intern - AI Systems and Infrastructure Engineering

Micron Technology • Austin (TX)

On-site
USD 34,000 - 44,000