Senior Software Engineer - VLM Microservices for Neural Reconstruction

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 152,000 - 241,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity options
Diversity and inclusion initiatives

Job summary

A leading tech company in California is seeking a highly skilled individual to develop and optimize containerized inference execution for their cutting-edge AI models. The role involves collaborating closely with Product and Research teams, enhancing model performance, and releasing production-grade software. Candidates should have advanced degrees in Computer Science or Electrical Engineering, coupled with extensive experience in AI systems, software engineering, and technologies like Docker and Kubernetes. Competitive salary based on experience and location.

Qualifications

  • 3+ years experience in AI distributed systems.
  • Proven track record in backend services and microservices.
  • Excellent software engineering fundamentals required.

Responsibilities

  • Design, build, and optimize inference execution for 3D VLMs.
  • Develop benchmarks for models' accuracy and performance.
  • Collaborate with Research and Product teams.

Skills

Docker
Kubernetes
Python
C++
CI/CD
Testing/Validation
Microservices
3D graphics
Machine Learning model engineering

Education

Master's of Science in Computer Science
Bachelor of Science in Electrical Engineering

Tools

vLLM
Torch
TRT
TRT-LLM
SGLang

Job description

NVIDIA is a world-leader in Gaussian Splatting and Neural reconstruction. Our team builds the Omniverse NuRec SDK to enable robotic, healthcare, and AV developers to build better models faster with closed-loop validation and closed-loop training grounded in real-world scenarios.**What you'll be doing:*** Design, build, and optimize containerized inference execution for the latest 3D VLMs from NVIDIA, turning research work into production-grade, highly optimized software (NIMs, NVIDIA Inference Microservices)* Develop benchmarks to validate the models accuracy and performance (latency, throughput, scalability)* Release and maintain the models and their pipelines throughout their lifecycle (bug fixes, security patches)* Contribute VLM-related features to Open-Source projects like vLLM* Collaborate closely with Research and Product teams and influence our common roadmaps**What we need to see:*** Master's of Science in Computer Science + 3 years, or Electrical Engineering, Bachelor of Science (or equivalent experience) + 5 years of experience.* History of building, validating and releasing production-grade AI distributed systems, backend services, microservices, and cloud technologies.* Deep technical expertise in distributed applications using Docker, Kubernetes, endpoints and their APIs (REST, gRPC), Helm.* Hands-on experience with modern inference platforms (vLLM, SGLang, Torch, TRT, TRT-LLM).* Proficiency with Python and C++.* Excellent software engineering fundamentals (source control, CI/CD, testing/validation, packaging, containerization).* Excellent written, visual, and verbal communication.* Curiosity and drive to learn new technologies and partner across teams and functions.**Ways to Stand Out from the Crowd:*** Experience with ML model engineering: training, fine-tuning, distillation, quantization.* Experience with low-level optimization of ML models (CUDA kernels)* Strong fundamentals in 3D graphics, 3D computer vision or neural reconstruction (NERFs, Gaussian Splats).* History of multidisciplinary creativity and innovation around software engineering in multiple problem domains.Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.You will also be eligible for equity and .Applications for this job will be accepted at least until February 16, 2026.This posting is for an existing vacancy.NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer - VLM Microservices for Neural Reconstruction
Senior Software Engineer - VLM Microservices for Neural Reconstruction

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior System Software Engineer - Neural Graphics SDKs
Senior System Software Engineer - Neural Graphics SDKs

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Corporation • Northern (KY)

On-site
USD 152,000 - 288,000
Senior Software Engineer, AI and DL Kernel Libraries
Senior Software Engineer, AI and DL Kernel Libraries

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior AI Software Engineer, Kernel Libraries
Senior AI Software Engineer, Kernel Libraries

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Principal Deep Learning Algorithm Engineer
Principal Deep Learning Algorithm Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Senior AI Workflow Engineer
Senior AI Workflow Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 184,000 - 357,000
Equity compensation
Health insurance
Relocation support
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Austin (TX)

On-site
USD 224,000 - 432,000
Equity
Benefits
Senior System Software Engineer - Neural Graphics SDKs
Senior System Software Engineer - Neural Graphics SDKs

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity
Benefits
Senior Software Engineer - GPU Local AI Platforms
Senior Software Engineer - GPU Local AI Platforms

NVIDIA • Westford (MA)

On-site
USD 224,000 - 432,000