Principal Deep Learning Algorithm Engineer

NVIDIA

United States

On-site

USD 272,000 - 431,000

Full time

12 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

United States Digital Space LLC is seeking outstanding engineers to advance LLM inference, improve performance, and develop scalable systems for agentic workflows. You will explore generative AI research and implement efficient inference algorithms within our software stack, collaborating across teams and partners to define system requirements.

Ideal candidates have 15+ years in deep learning systems, strong Python and C++ skills, and a solid grasp of computer architecture and GPU-based

Qualifications

  • BS/MS/PhD in CS/EE/CE or related field (or equivalent experience).
  • 15+ years of experience in deep learning and deep learning systems design.
  • Proficiency in Python and C++ programming.
  • Strong understanding of computer architecture, and GPU/parallel datacenter computing fundamentals.

Responsibilities

  • Research and Development: Explore contemporary research on generative AI, agents, and inference systems into the LLM software stack.
  • Workload Analysis and Optimization: Profile and optimize agentic LLM workloads to reduce latency and increase throughput.
  • System Design and Implementation: Design scalable systems to accelerate agentic workflows and handle datacenter-scale use cases.
  • Collaboration and Communication: Advise on software/hardware/system iterations by engaging with teams and external partners.

Skills

Deep learning
Python
C++
GPU computing
Performance modeling

Education

BS/MS/PhD in CS/EE/CE or related field

Tools

CUDA
OpenCL

Job description

At the company, we are at the forefront of the constantly evolving field of large language models, and their application in agentic and reasoning use cases. As the scale and complexity of these LLM systems continues to increase, we are seeking outstanding engineers to join our team and help shape the future of LLM inference.

Our team is dedicated to pushing the boundaries of what's possible with LLMs by improving the algorithmic performance and efficiency of systems that represent them. We constantly reflect on how to improve these systems, developing new inference algorithms and protocols, improving existing models, and seamlessly integrating improvements to ensure the company's solutions can efficiently handle large-scale, sophisticated tasks.

What you’ll be doing:
  • Research and Development: Explore and incorporate contemporary research on generative AI, agents, and inference systems into the the company LLM software stack.
  • Workload Analysis and Optimization: Conduct in-depth analysis, profiling, and optimization of agentic LLM workloads to significantly reduce request latency and increase request throughput while maintaining workflow fidelity.
  • System Design and Implementation: Design and implement scalable systems to accelerate agentic workflows and efficiently handle sophisticated datacenter-scale use cases.
  • Collaboration and Communication: Advise future iterations of the company software, hardware, and system by engaging with a diverse set of teams at the company and external partners and formalizing the strategic requirements presented by their workloads.
What we need to see:
  • BS, MS, PhD in Computer Science, Electrical Engineering, Computer Engineering, or a related field (or equivalent experience).
  • 15+ years of experience in deep learning and deep learning systems design.
  • Proficiency in Python and C++ programming
  • Strong understanding of computer architecture, and GPU/parallel datacenter computing fundamentals.
  • Proven interest in analyzing, modeling, and tuning application performance.
Ways to stand out from the crowd:
  • Experience in building large-scale LLM inference systems, especially those involving compound AI.
  • Experience with processor and system-level performance modeling.
  • GPU programming experience with CUDA or OpenCL.

the company has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 26, 2026.This posting is for an existing vacancy.the company uses AI tools in its recruiting processes.

the company is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.Originally posted on Himalayas

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior System Software Engineer, Agentic Retrieval
Senior System Software Engineer, Agentic Retrieval

NVIDIA • United States

Remote
USD 184,000 - 357,000
Equity
Benefits
Machine Learning Engineer
Machine Learning Engineer

Thomas To • Santa Clara (CA)

On-site
USD 170,000 - 242,000
Equity
Benefits
Senior System Software Engineer, Agentic Retrieval
Senior System Software Engineer, Agentic Retrieval

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Principal LLM Application Engineer
Principal LLM Application Engineer

system_two • Palo Alto (CA)

Remote
USD 180,000 - 260,000
Medical, Dental, Vision
Equity in the company
Software Engineer, Model Inference, DeepMind
Software Engineer, Model Inference, DeepMind

DeepMind Technologies Limited • Madison (WI)

On-site
USD 207,000 - 300,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Machine Learning Engineer
Machine Learning Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity
Benefits package
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA • Washington

On-site
USD 184,000 - 288,000
Equity
Benefits
Principal High-Performance LLM Training Engineer
Principal High-Performance LLM Training Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits