Senior GPU & AI Systems Engineer - Multi-Node Training

Micron Memory Malaysia Sdn Bhd

Boise (ID)

On-site

USD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical benefits
Paid time off
401(k) plan

Job summary

Micron Technology's Boise, Idaho team seeks a seasoned GPU Performance Engineer to architect and optimize large-scale AI workloads on multi-GPU clusters. You'll train, fine-tune, and deploy models (SFT, RLHF) and push performance through distributed training and memory-efficiency techniques.

You will collaborate with data scientists, hardware architects, and software engineers, mentoring peers while advancing CUDA/HIP kernels, ML frameworks, and GenAI agent strategies.

Qualifications

  • 10+ years in GPU hardware and software optimization across cloud and on‑prem.
  • 5+ years in performance optimization and low‑level systems using C++ and CUDA.
  • Deep proficiency with LLMs, prompt tooling, and PEFT methods (LoRA/QLoRA).
  • Experience building scalable ML systems with distributed training (DDP, FSDP).
  • Proficient in Python (preferred) or Java and CI/CD with Docker/Kubernetes.

Responsibilities

  • Architect and execute large-scale model training and fine-tuning on multi-node, multi-GPU clusters.
  • Optimize throughput and memory using distributed strategies and mixed precision.
  • Design autonomous AI agents for multi-step reasoning and tool execution.
  • Profile workloads to identify bottlenecks in compute, memory, and latency.
  • Write high-performance kernels (CUDA/HIP) and collaborate with hardware architects.
  • Mentor engineers in parallel programming and optimization techniques.

Skills

GPU architecture
Performance optimization
Distributed training
Python
C++
LLMs and PEFT
GenAI / AI agents
CI/CD tooling
Docker & Kubernetes
Communication skills

Education

Bachelor’s or Master’s in CS/Statistics or related
PhD preferred

Tools

CUDA
Triton
PyTorch
Kubernetes
Ray
Kubeflow
TensorRT-LLM

Job description

Micron Technology's Boise, Idaho team seeks a seasoned GPU Performance Engineer to architect and optimize large-scale AI workloads on multi-GPU clusters. You'll train, fine-tune, and deploy models (SFT, RLHF) and push performance through distributed training and memory-efficiency techniques.

You will collaborate with data scientists, hardware architects, and software engineers, mentoring peers while advancing CUDA/HIP kernels, ML frameworks, and GenAI agent strategies.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff AI Engineer — GPU Performance & GenAI
Staff AI Engineer — GPU Performance & GenAI

Micron Technology, Inc • Boise (ID)

On-site
USD 120,000 - 150,000
Medical, dental, and vision plans
Paid time off and holidays
Income protection programs
Senior AI Systems Architect — Multinode GPU ML
Senior AI Systems Architect — Multinode GPU ML

Micron Technology, Inc • Boise (ID)

On-site
USD 130,000 - 170,000
Medical, dental, and vision plans
Paid family leave
Robust paid time-off
Senior AI Systems Engineer: GPU ML & Multi-Agent Automation
Senior AI Systems Engineer: GPU ML & Multi-Agent Automation

1000 Micron Technology, Inc. • Boise (ID)

On-site
USD 120,000 - 150,000
Medical, dental, and vision plans
Paid family leave
Robust paid time-off program
AI/ML Engineer & Architect: Lead Scalable AI Solutions
AI/ML Engineer & Architect: Lead Scalable AI Solutions

Micron Technology, Inc • Boise (ID)

On-site
USD 140,000 - 210,000
Senior AI Training Performance Engineer (GPU & Scale)
Senior AI Training Performance Engineer (GPU & Scale)

figure.ai • San Jose (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
Lead GPU ML Training Performance Engineer
Lead GPU ML Training Performance Engineer

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 120,000 - 160,000
Senior GPU/AI Systems Engineer - Performance & ML
Senior GPU/AI Systems Engineer - Performance & ML

AMD • Santa Clara (CA)

On-site
USD 170,000 - 250,000
AMD Benefits
Senior ML Training Systems Engineer - Distributed CUDA
Senior ML Training Systems Engineer - Distributed CUDA

Genesis AI • San Francisco (CA)

On-site
USD 180,000 - 260,000
Member of Technical Staff, AI Engineering
Member of Technical Staff, AI Engineering

Micron Technology • Boise (ID)

On-site
USD 130,000 - 170,000
Medical, dental, and vision plans
Paid family leave
Robust paid time-off
Staff Principal AI Engineer - Full-Stack Innovator
Staff Principal AI Engineer - Full-Stack Innovator

1000 Micron Technology, Inc. • Boise (ID)

On-site
USD 140,000 - 190,000
Medical, dental, vision plans
Paid time off & holidays
Paid family leave