AI Engineer — LLM / VLM

SAI Group Ltd

Kolkata District

On-site

INR 1,500,000 - 3,500,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

SAI Group Ltd is seeking an AI Engineer specializing in Large Language Models (LLMs) and Vision-Language Models (VLMs) to design, develop, and deploy production-grade AI solutions in Kolkata. The role focuses on prompt engineering, RAG, fine-tuning, multimodal AI, and model serving.

The candidate should have hands-on experience with LLMs/VLMs, Transformers, PyTorch, and Hugging Face, and be able to build production-grade APIs and pipelines using Python, FastAPI, Docker, and cloud platforms.

Qualifications

  • Strong Python programming and software-engineering fundamentals.
  • Hands-on experience with LLMs and/or VLMs.
  • Strong understanding of Transformers, attention mechanisms, tokenization, embeddings, and inference.
  • Experience with PyTorch and Hugging Face Transformers.
  • Experience building RAG systems and vector-search solutions.
  • Knowledge of prompt engineering and LLM evaluation.
  • Experience with APIs, REST services, Git, Docker, and CI/CD.
  • Familiarity with vector databases such as FAISS, Milvus, Pinecone, Weaviate, or pgvector.
  • Understanding of cloud AI infrastructure, preferably AWS/Azure/GCP.

Responsibilities

  • Develop and deploy AI applications using LLMs and VLMs.
  • Build RAG pipelines involving document ingestion, chunking, embeddings, retrieval, reranking, and generation.
  • Work with models such as GPT, Claude, Gemini, Llama, Mistral, Qwen, and multimodal/VLM models.
  • Develop multimodal solutions involving text, images, PDFs, charts, tables, and documents.
  • Perform prompt engineering, supervised fine-tuning, LoRA/QLoRA, and model evaluation.
  • Build AI agents and tool-calling workflows where appropriate.
  • Optimize inference for latency, throughput, memory, and cost.
  • Develop APIs and production services using Python, FastAPI, Docker, and cloud platforms.
  • Implement evaluation frameworks to measure accuracy, hallucination, relevance, latency, and safety.
  • Collaborate with ML engineers, software engineers, and product teams to take prototypes into production.

Skills

Python programming
LLMs/VLMs experience
Transformers/attention
Prompt engineering
RAG pipelines
APIs REST
CI/CD
Cloud platforms (AWS/Azure/GCP)
Vector databases knowledge

Tools

Docker
Git
REST APIs
CI/CD tooling
PyTorch
HuggingFace Transformers
FastAPI
FAISS
Milvus
Pinecone
Weaviate
pgvector

Job description

Role Overview

We are looking for anAI Engineer specializing in Large Language Models (LLMs) and Vision-Language Models (VLMs)to design, develop, and deploy production-grade AI solutions. The ideal candidate should have strong experience with LLM/VLM architectures, prompt engineering, RAG, fine-tuning, multimodal AI, and model serving.

Key Responsibilities
  • Develop and deploy AI applications using LLMs and VLMs.
  • Build RAG pipelines involving document ingestion, chunking, embeddings, retrieval, reranking, and generation.
  • Work with models such as GPT, Claude, Gemini, Llama, Mistral, Qwen, and multimodal/VLM models.
  • Develop multimodal solutions involving text, images, PDFs, charts, tables, and documents.
  • Perform prompt engineering, supervised fine-tuning, LoRA/QLoRA, and model evaluation.
  • Build AI agents and tool-calling workflows where appropriate.
  • Optimize inference for latency, throughput, memory, and cost.
  • Develop APIs and production services using Python, FastAPI, Docker, and cloud platforms.
  • Implement evaluation frameworks to measure accuracy, hallucination, relevance, latency, and safety.
  • Collaborate with ML engineers, software engineers, and product teams to take prototypes into production.
Required Skills
  • Strong Python programming and software-engineering fundamentals.
  • Hands-on experience with LLMs and/or VLMs.
  • Strong understanding of Transformers, attention mechanisms, tokenization, embeddings, and inference.
  • Experience with PyTorch and Hugging Face Transformers.
  • Experience building RAG systems and vector-search solutions.
  • Knowledge of prompt engineering and LLM evaluation.
  • Experience with APIs, REST services, Git, Docker, and CI/CD.
  • Familiarity with vector databases such as FAISS, Milvus, Pinecone, Weaviate, or pgvector.
  • Understanding of cloud AI infrastructure, preferably AWS/Azure/GCP.
VLM / Computer Vision Skills
  • Experience with multimodal models such as Qwen-VL, LLaVA, Gemini, GPT vision models, or similar.
  • Understanding of image preprocessing and document/image understanding.
  • Experience with OCR, document intelligence, image classification, object detection, or visual question answeringis a plus.
  • Ability to build pipelines combining vision + language + retrieval.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Engineer (LLMs, Agentic Systems & Model Training)
AI Engineer (LLMs, Agentic Systems & Model Training)

Kayana | Ordering & Payment Solutions • Mumbai

On-site
INR 1,200,000 - 2,000,000
Competitive salary and benefits
Opportunity to work with cutting-edge AI systems
Collaborative environment
+1
AI Engineer
AI Engineer

Keka Inc. • Bengaluru

On-site
INR 1,200,000 - 1,800,000
AI Developer - Large Language Models
AI Developer - Large Language Models

Volody • Mumbai

On-site
INR 1,800,000 - 2,400,000
AI Software Engineer-1
AI Software Engineer-1

Detect Technologies • Chennai District

On-site
INR 600,000 - 1,200,000
AI Developer - Large Language Models
AI Developer - Large Language Models

Volody • Maharashtra

On-site
INR 1,200,000 - 1,800,000
AI/ML Engineer
AI/ML Engineer

Bacancy Technology Inc • Ahmedabad District

On-site
INR 600,000 - 900,000
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Bhavitha Tech, CMMi Level 3 Company • Bengaluru

On-site
INR 1,200,000 - 1,800,000
AI Engineer
AI Engineer

Biz4Group LLC • Jaipur

On-site
INR 1,500,000 - 2,500,000
AI Engineer (LLM / Chatbot / RAG)
AI Engineer (LLM / Chatbot / RAG)

Qentelli • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Vision AI Solution Architect
Vision AI Solution Architect

Tata Consultancy Services • Bengaluru

On-site
INR 3,000,000 - 5,400,000