Senior AI Software Engineer, Kernel Libraries

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 184,000 - 287,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading technology company is seeking a Senior AI Software Engineer to join their team in Santa Clara, California. In this role, you will innovate and develop groundbreaking AI systems software for inference applications including deep learning framework optimizations and GPU kernel technologies. You will closely collaborate with other engineers to enhance NVIDIA’s leading technology and ensure the efficiency of AI workloads. Ideal candidates will possess a Master's degree in Computer Science, strong programming skills in C/C++, and significant experience with ML frameworks.

Qualifications

  • 6+ years of experience in ML/DL systems development preferred.
  • Strong experience with GPU kernel development.
  • PhD preferred for advanced roles.

Responsibilities

  • Innovate and develop new AI systems technologies.
  • Design and optimize kernels for AI workloads.
  • Collaborate with other engineers on deep learning frameworks.

Skills

Deep learning frameworks (e.g. PyTorch, JAX, TensorFlow, ONNX)
C/C++ programming
Machine Learning/Deep Learning systems development

Education

Master's degree in Computer Science or Electrical Engineering

Tools

CUDA C/C++
Apache TVM
MLIR

Job description

Senior AI Software Engineer, Kernel Libraries page is loaded## Senior AI Software Engineer, Kernel Librarieslocations: US, CA, Santa Clara: US, Remotetime type: Full timeposted on: Posted Yesterdayjob requisition id: JR2014705We're looking for outstanding AI systems engineers to develop groundbreaking technologies in the inference systems software stack! We build innovative AI systems software to accelerate for AI inference. As a member of the team, you'll develop libraries, code generators, and GPU kernel technologies for NVIDIA's hardware architecture. This means designing and building things like new abstractions, efficient attention kernel implementations, new LLM inference runtimes components, and kernel code generators to accelerate large language models, agents, and other high-impact AI workloads.**What you'll be doing:*** Innovating and developing new AI systems technologies for efficient inference* Designing, implementing, and optimizing kernels for high impact AI workloads* Designing and implementing extensible abstractions for LLM serving engines* Building efficient just-in-time domain specific compilers and runtimes* Collaborating closely with other engineers at NVIDIA across deep learning frameworks, libraries, kernels, and GPU arch teams* Contributing to open source communities like FlashInfer, vLLM, and SGLang**What we need to see:*** Masters degree in Computer Science, Electrical Engineering, or related field (or equivalent experience); PhD are preferred* 6+ years (academic/ industry) experience with ML/DL systems development preferable* Strong experience in developing or using deep learning frameworks (e.g. PyTorch, JAX, TensorFlow, ONNX, etc) and ideally inference engines and runtimes such as vLLM, SGLang, and MLC.* Strong Python and C/C++ programming skills**Ways to stand out from the crowd:*** Background in domain specific compiler and library solutions for LLM inference and training (e.g. FlashInfer, Flash Attention)* Expertise in inference engines like vLLM and SGLang* Expertise in machine learning compilers (e.g. Apache TVM, MLIR)* Strong experience in GPU kernel development and performance optimizations (especially using CUDA C/C++, cuTile, Triton, or similar)* Open source project ownership or contributionsYour base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.You will also be eligible for equity and .Applications for this job will be accepted at least until March 15, 2026.This posting is for an existing vacancy.NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer, AI and DL Kernel Libraries
Senior Software Engineer, AI and DL Kernel Libraries

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Corporation • Northern (KY)

Hybrid
USD 152,000 - 288,000
Senior Software Engineer, Machine Learning Inference
Senior Software Engineer, Machine Learning Inference

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 152,000 - 288,000
Equity
Benefits
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Corporation • California (MO)

On-site
USD 152,000 - 287,500
Equity
Benefits
Senior DL Algorithms Engineer - Inference Performance
Senior DL Algorithms Engineer - Inference Performance

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 184,000 - 288,000
Senior Software Engineer, AI Storage
Senior Software Engineer, AI Storage

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity options
Competitive salary based on experience
Principal Deep Learning Algorithm Engineer
Principal Deep Learning Algorithm Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Senior Infrastructure Software Engineer, Deep Learning Libraries
Senior Infrastructure Software Engineer, Deep Learning Libraries

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior AI Workflow Engineer
Senior AI Workflow Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 184,000 - 357,000
Equity compensation
Health insurance
Relocation support
Senior Compiler Engineer, AI Inference Platforms
Senior Compiler Engineer, AI Inference Platforms

NVIDIA Corporation • Austin (TX)

On-site
USD 152,000 - 242,000
Competitive salary
Generous benefits package
Equity options