Senior LLM Efficiency Architect: Model Systems Co-Design

Nvidia Corporation in

Santa Clara (CA)

Hybrid

USD 184,000 - 357,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in Santa Clara, CA is seeking a Senior DL Performance Efficiency Architect (Finance) to lead the strategy and execution for making large language models more efficient from research to deployment. You will drive cross-layer optimization, collaborate with researchers and hardware teams, and shape how future models and platforms are designed for practical compute, memory, power, and cost constraints.

The role requires hands-on leadership, deep knowledge of LLM workloads, and a track record

Qualifications

  • MS or PhD in CS/EE/CE or equivalent experience.
  • 5+ years in AI systems, model architecture, or performance optimization.
  • Strong understanding of LLM architectures, training and inference workloads.
  • Proven ability to lead complex optimization projects from concept to production.
  • Experience with roofline modeling, benchmarking and hardware-aware optimization.

Responsibilities

  • Lead cross-layer efforts to improve the efficiency of large language models across model architecture, training and inference systems.
  • Analyze how LLM workloads map to GPUs, memory systems, interconnects, and distributed infrastructure, and identify opportunities for model-system-hardware co-design.
  • Establish a measurement-driven efficiency roadmap and lead projects from early investigation through production deployment.
  • Partner with model researchers, systems engineers, compiler and kernel developers, and hardware architects to influence future model, software, and hardware roadmaps.

Skills

AI systems
Model architecture
Computer architecture
Performance optimization
High-performance computing

Education

MS or PhD in CS/EE/CE or equivalent

Job description

NVIDIA in Santa Clara, CA is seeking a Senior DL Performance Efficiency Architect (Finance) to lead the strategy and execution for making large language models more efficient from research to deployment. You will drive cross-layer optimization, collaborate with researchers and hardware teams, and shape how future models and platforms are designed for practical compute, memory, power, and cost constraints.

The role requires hands-on leadership, deep knowledge of LLM workloads, and a track record

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior LLM Efficiency Architect — Model-System Optimizer
Senior LLM Efficiency Architect — Model-System Optimizer

NVIDIA • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
LLM Efficiency Architect: Cross-Layer Performance Leader
LLM Efficiency Architect: Cross-Layer Performance Leader

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Senior DL Performance Efficiency Architect
Senior DL Performance Efficiency Architect

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Senior DL Performance Efficiency Architect
Senior DL Performance Efficiency Architect

NVIDIA • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Senior DL Performance Efficiency Architect
Senior DL Performance Efficiency Architect

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Senior LLM Inference Engineer: Performance & Optimization
Senior LLM Inference Engineer: Performance & Optimization

Confidential • United States

On-site
USD 180,000 - 240,000
Senior LLM Training Performance Architect
Senior LLM Training Performance Architect

NVIDIA • Santa Clara (CA)

Hybrid
USD 184,000 - 356,500
Equity
Benefits
Lead Architect, High-Performance LLM Training
Lead Architect, High-Performance LLM Training

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Senior LLM Infra Engineer — AI Model Serving
Senior LLM Infra Engineer — AI Model Serving

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits package
Competitive salary
Senior LLM Inference Algorithms Engineer — Equity Options
Senior LLM Inference Algorithms Engineer — Equity Options

NVIDIA • California (MO)

On-site
USD 272,000 - 432,000
Equity
Benefits