Principal AI Memory Architect

Conductor

San Jose (CA)

On-site

USD 219,000 - 351,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Medical/Dental/Vision/401k
4+ weeks paid time off
Fertility/adoption support

Job summary

Samsung Semiconductor in San Jose seeks a Hands-on Principal Engineer to own memory footprint and inference performance for next-generation AI models. Define tiered memory architectures from HBM to NVMe, translating model behavior into hardware and software requirements.

You will lead architecture reviews, mentor engineers, and represent the company in CTO-to-CTO discussions, requiring deep knowledge of MoE, transformer internals, and production memory management across stacks like vLLM,

Qualifications

  • First-principles understanding of transformer-class model internals.
  • Working knowledge of MoE model behavior and routing.
  • Experience profiling memory bottlenecks in production serving.

Responsibilities

  • Define memory architecture requirements for tiered inference at fleet scale.
  • Lead architecture reviews and cross-team design sessions.
  • Mentor senior engineers and contribute to technical narratives for customers.

Skills

Transformer internals
MoE model behavior
Memory‑system design
Performance engineering

Education

BS in Computer/Electrical/CS
MS in Computer/Electrical/CS

Tools

vLLM PagedAttention
SGLang HiCache
TensorRT-LLM
llama.cpp

Job description

Samsung Semiconductor in San Jose seeks a Hands-on Principal Engineer to own memory footprint and inference performance for next-generation AI models. Define tiered memory architectures from HBM to NVMe, translating model behavior into hardware and software requirements.

You will lead architecture reviews, mentor engineers, and represent the company in CTO-to-CTO discussions, requiring deep knowledge of MoE, transformer internals, and production memory management across stacks like vLLM,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior RTL Engineer: Memory-Centric AI/ML IP
Senior RTL Engineer: Memory-Centric AI/ML IP

Socket.dev • San Jose (CA)

On-site
USD 138,000 - 206,000
Charitable giving match
Paid time off 4+ weeks
Fertility/adoption stipend
+3
Senior AI System Architect for Memory & Compute
Senior AI System Architect for Memory & Compute

Sandisk • Milpitas (CA)

On-site
USD 170,000 - 260,000
Senior Staff Engineer — AI Inference Co-Design Lead
Senior Staff Engineer — AI Inference Co-Design Lead

Samsung Semiconductor • San Jose (CA)

Hybrid
USD 189,000 - 301,000
Medical/Dental/Vision
401k
Wellness apps
+2
Senior AI System Architect - Memory & Inference Lead
Senior AI System Architect - Memory & Inference Lead

Sandisk • California (MO)

On-site
USD 140,000 - 190,000
Medical insurance
Dental insurance
401(k) plan
+1
Memory Architect: AI Hardware Strategy & Optimization
Memory Architect: AI Hardware Strategy & Optimization

HP • Spring (TX)

On-site
USD 140,000 - 190,000
Health insurance
Dental insurance
Vision insurance
+7
Senior Staff Engineer, AI Workloads & Storage Architect
Senior Staff Engineer, AI Workloads & Storage Architect

Conductor • San Jose (CA)

On-site
USD 189,000 - 301,000
4+ weeks of paid time off
Medical/Dental/Vision/401k
Charitable giving match
AI Memory Sales Lead – Global Revenue Strategy
AI Memory Sales Lead – Global Revenue Strategy

Samsung Semiconductor Inc. • Bellevue (WA)

On-site
USD 158,000 - 252,000
Medical/Dental/Vision/401k
Incentive opportunities
Paid time off: 4+ weeks + holidays +  
+1
Senior RTL IP Engineer - Memory-Centric AI/ML
Senior RTL IP Engineer - Memory-Centric AI/ML

Samsung Semiconductor • San Jose (CA)

On-site
USD 138,000 - 206,000
4+ weeks PTO
Onsite Café & gym
Fertility care stipend
+4
Lead Architect - CPU Architecture & Performance
Lead Architect - CPU Architecture & Performance

Samsungsemiconductor • San Jose (CA)

On-site
USD 219,000 - 351,000
Charitable giving match
Paid time off 4+ weeks
Fertility/adoption support
+1
Senior Memory Qualification Engineer — AI-Driven, Equity
Senior Memory Qualification Engineer — AI-Driven, Equity

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 311,000
Equity
Benefits
Flexible time off
+1