Senior Solutions Architect, Generative AI

NVIDIA Corporation

Hinoba-an

On-site

PHP 2,400,000 - 3,600,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

NVIDIA Corporation is seeking a dynamic Generative AI Solution Architect with deep expertise in training Large Language Models (LLMs) and Agentic AI. You will architect end-to-end generative AI solutions, collaborate with customers to tackle language-focused challenges, and work with sales to showcase LLM/RAG capabilities.

You will lead workshops, train LLMs on NVIDIA’s hardware and software platforms, and guide deployment for optimized inference.

Qualifications

  • 10+ years hands-on experience in a technical role focusing on generative AI, with emphasis on training LLMs.
  • Proven track record deploying and optimizing LLM models for inference in production.
  • Deep understanding of state-of-the-art language models (e.g., GPT-3, BERT).
  • Expertise in training and fine-tuning LLMs using TensorFlow, PyTorch, or Hugging Face Transformers.
  • Proficiency in model deployment/optimization for efficient GPU inference.
  • Strong communication and collaboration skills with stakeholders.
  • Experience leading workshops and presenting technical solutions.

Responsibilities

  • Architect end-to-end generative AI solutions focusing on LLMs and RAG workflows.
  • Collaborate with customers to understand language-related business challenges and design tailored solutions.
  • Support pre-sales with technical presentations and demonstrations of LLM/RAG capabilities.
  • Collaborate with NVIDIA engineering teams to provide feedback and contribute to technology evolution.
  • Lead workshops and design sessions to define generative AI solutions and train/optimize LLMs on NVIDIA platforms.
  • Design and implement RAG-based workflows to enhance content generation and retrieval.
  • Guide customers in integrating RAG workflows into their applications and systems.

Skills

Generative AI
LLM training
Leadership
Communication

Education

B.Tech, M.S. or Ph.D. in Computer Science/Artificial Intelligence

Tools

TensorFlow
PyTorch
Hugging Face Transformers
Docker
Kubernetes

Job description

NVIDIA is seeking a dynamic and experienced Generative AI Solution Architect with specialized expertise in training Large Language Models (LLMs) and Agentic AI. As a key member of our AI Solutions team, you will play a pivotal role in architecting and delivering cutting-edge solutions that leverage the power of NVIDIA's generative AI technologies.

What you will be doing:
  • Architect end-to-end generative AI solutions with a focus on LLMs, Agentic and RAG workflows.
  • Collaborate closely with customers to understand their language-related business challenges and design tailored solutions.
  • Collaborate with sales and business development teams to support pre-sales activities, including technical presentations and demonstrations of LLM and RAG capabilities.
  • Work closely with NVIDIA engineering teams to provide feedback and contribute to the evolution of generative AI technologies.
  • Engage directly with customers to understand their language-related requirements and challenges.
  • Lead workshops and design sessions to define and refine generative AI solutions focused on LLMs and RAG workflows and lead the training and optimization of Large Language Models using NVIDIA’s hardware and software platforms.
  • Implement strategies for efficient and effective training of LLMs to achieve optimal performance.
  • Design and implement RAG-based workflows to enhance content generation and information retrieval.
  • Work closely with customers to integrate RAG workflows into their applications and systems and stay abreast of the latest developments in language models and generative AI technologies.
  • Provide technical leadership and guidance on best practices for training LLMs and implementing RAG-based solutions.
What we need to see:
  • B.Tech ,Master's or Ph.D. in Computer Science, Artificial Intelligence, or equivalent experience
  • 10+ years of hands-on experience in a technical role, specifically focusing on generative AI, with a strong emphasis on training Large Language Models (LLMs).
  • Proven track record of successfully deploying and optimizing LLM models for inference in production environments.
  • In-depth understanding of state-of-the-art language models, including but not limited to GPT-3, BERT, or similar architectures.
  • Expertise in training and fine-tuning LLMs using popular frameworks such as TensorFlow, PyTorch, or Hugging Face Transformers.
  • Proficiency in model deployment and optimization techniques for efficient inference on various hardware platforms, with a focus on GPUs.
  • Strong knowledge of GPU cluster architecture and the ability to leverage parallel processing for accelerated model training and inference.
  • Excellent communication and collaboration skills with the ability to articulate complex technical concepts to both technical and non-technical stakeholders.
  • Experience leading workshops, training sessions, and presenting technical solutions to diverse audiences.
Ways to stand out from the crowd:
  • Proven ability to optimize LLM models for inference speed, memory efficiency, and resource utilization.
  • Familiarity with containerization technologies (e.g., Docker) and orchestration tools (e.g., Kubernetes) for scalable and efficient model deployment.
  • Deep understanding of GPU cluster architecture, parallel computing, and distributed computing concepts.
  • Hands-on experience with NVIDIA GPU technologies, and GPU cluster management and ability to design and implement scalable and efficient workflows for LLM training and inference on GPU clusters.

With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing.

If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you!

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior GenAI Solutions Architect: LLM & RAG Expert
Senior GenAI Solutions Architect: LLM & RAG Expert

NVIDIA Corporation • Hinoba-an

On-site
PHP 2,400,000 - 3,600,000
Generative AI Architect/ Lead Data Scientist
Generative AI Architect/ Lead Data Scientist

V2 Solutions • Hinoba-an

On-site
PHP 1,200,000 - 2,000,000
Director Associate Distinguished Engineer (Solution Architect - GenAI + RAG + Agentic AI)
Director Associate Distinguished Engineer (Solution Architect - GenAI + RAG + Agentic AI)

Nagarro • Hinoba-an

On-site
PHP 2,400,000 - 4,200,000
Lead Data Scientist
Lead Data Scientist

V2 Solutions • Hinoba-an

On-site
PHP 1,200,000 - 1,800,000
IN_Sr Associate__Generative AI Engineer
IN_Sr Associate__Generative AI Engineer

V2 Solutions • Hinoba-an

On-site
PHP 900,000 - 1,300,000
IN_Sr Associate__Generative AI Engineer_Advisory_Bangalore
IN_Sr Associate__Generative AI Engineer_Advisory_Bangalore

V2 Solutions • Hinoba-an

On-site
PHP 1,674,000 - 3,571,000
Lead Architect
Lead Architect

V2 Solutions • Hinoba-an

Hybrid
PHP 1,194,000 - 2,122,000
Senior Software Engineering Lead
Senior Software Engineering Lead

Vanigent Biopharm • Metro Manila

On-site
PHP 3,000,000 - 5,000,000
AI Engineer
AI Engineer

Dry Ground • Philippines

On-site
PHP 2,995,000 - 4,794,000
Competitive salary and performance-based incentives
Flexible work environment
Collaborative innovation-driven culture
GenAI Engineer - Hybrid Role with Growth & Benefits
GenAI Engineer - Hybrid Role with Growth & Benefits

Thinking Machines Data Science • Philippines

Hybrid
PHP 800,000 - 1,200,000
Hybrid Set-Up
Health benefits
Professional development budget
+1