LLM Engineer (Large Language Models)

Fospe UK Ltd

Bengaluru

Hybrid

INR 2,500,000 - 5,200,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive compensation with bonuses
Hybrid work at Bangalore Innovation Cn
Health, dental, wellness insurance
AI research learning stipend
MacBook Pro + equipment budget
Stock options

Job summary

Fospe is seeking an LLM Engineer to design, fine-tune, and operationalize cutting-edge Large Language Models and RAG pipelines for enterprise SaaS products.

You will collaborate with our AI research team to embed deterministic cognitive capabilities and high-throughput semantic reasoning, building scalable inference pipelines and guardrail systems.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or related quantitative field.
  • 3+ years of professional experience in machine learning and deep learning, with 1+ years dedicated to LLM engineering.
  • Deep hands-on expertise with PyTorch, Hugging Face Transformers, vLLM, and LangChain/LlamaIndex.
  • Proven experience with vector databases such as Pinecone, Qdrant, Milvus, or pgvector.
  • Strong proficiency in Python, asynchronous programming, Docker containerization, and Linux/CUDA environments.
  • Demonstrated understanding of model optimization techniques including LoRA, QLoRA, AWQ, and GPTQ quantization.

Responsibilities

  • Architect, fine-tune, and quantize open-source foundation models for domain-specific enterprise workloads.
  • Design high-accuracy Retrieval-Augmented Generation (RAG) architectures with hybrid dense/sparse search, reranking, and semantic chunking.
  • Build low-latency streaming inference pipelines utilizing vLLM, TensorRT-LLM, and Triton Inference Server.
  • Implement guardrail validation, safety filters, prompt routing, and automated hallucination benchmarking.
  • Collaborate with backend engineers to expose scalable gRPC and REST interfaces for seamless frontend consumption.

Skills

LLM engineering
Python
asynchronous programming
Docker containerization
Linux/CUDA environments
Pytorch
Hugging Face Transformers
vLLM
LangChain/LlamaIndex
vector databases
Qdrant
Pinecone
Milvus
pgvector
Kubernetes
CUDA optimizations

Education

Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or related quantitative field

Tools

Docker
Kubernetes

Job description

About the Role

As an LLM Engineer at Fospe, you will be at the forefront of designing, fine-tuning, and operationalizing cutting-edge Large Language Models and Retrieval-Augmented Generation (RAG) pipelines. You will collaborate directly with our AI research team to embed deterministic cognitive capabilities and high-throughput semantic reasoning into our vertical enterprise SaaS products.

Key Responsibilities
  • Architect, fine-tune, and quantize open-source foundation models (Llama 3, Mistral, Qwen, DeepSeek) for domain-specific enterprise workloads.
  • Design high-accuracy Retrieval-Augmented Generation (RAG) architectures with hybrid dense/sparse search, reranking, and semantic chunking.
  • Build low-latency streaming inference pipelines utilizing vLLM, TensorRT-LLM, and Triton Inference Server.
  • Implement guardrail validation, safety filters, prompt routing, and automated hallucination benchmarking.
  • Collaborate with backend engineers to expose scalable gRPC and REST interfaces for seamless frontend consumption.
Requirements & Qualifications
  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or related quantitative field.
  • 3+ years of professional experience in machine learning and deep learning, with 1+ years dedicated to LLM engineering.
  • Deep hands-on expertise with PyTorch, Hugging Face Transformers, vLLM, and LangChain/LlamaIndex.
  • Proven experience with vector databases such as Pinecone, Qdrant, Milvus, or pgvector.
  • Strong proficiency in Python, asynchronous programming, Docker containerization, and Linux/CUDA environments.
  • Demonstrated understanding of model optimization techniques including LoRA, QLoRA, AWQ, and GPTQ quantization.
Preferred Qualifications
  • Experience with multi-agent orchestration frameworks (LangGraph, CrewAI, AutoGen).
  • Contributions to open-source AI libraries or published research in NLP/LLMs.
  • Experience deploying models on Kubernetes and cloud infrastructure (AWS/GCP/Azure GPU instances).
Primary Technologies

Python, PyTorch, Hugging Face, vLLM, LangChain, LlamaIndex, Qdrant, Docker, Kubernetes, CUDA, FastAPI

What We Offer
  • Competitive compensation package with performance bonuses
  • Flexible hybrid working schedule (Bangalore Innovation Center)
  • Comprehensive health, dental, and wellness insurance coverage
  • Generous annual compute and AI research learning stipend
  • Latest Apple MacBook Pro M-series workstation & equipment budget
  • Stock option and co-ownership opportunities
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Engineer
Machine Learning Engineer

Tranzeal • Bengaluru

On-site
INR 3,500,000 - 7,500,000
Gen AI Engineer
Gen AI Engineer

HCLTech • Bengaluru

On-site
INR 2,500,000 - 4,500,000
LLM Specialist
LLM Specialist

TRDFIN Support Services Pvt Ltd • Gurugram District

On-site
INR 1,000,000 - 1,500,000
AI Engineer (LLMs, Agentic Systems & Model Training)
AI Engineer (LLMs, Agentic Systems & Model Training)

Kayana | Ordering & Payment Solutions • Mumbai

On-site
INR 1,200,000 - 2,000,000
Competitive salary and benefits
Opportunity to work with cutting-edge AI systems
Collaborative environment
+1
Senior AI / LLM Engineer
Senior AI / LLM Engineer

Thinkscoop Technologies • Bengaluru

Hybrid
INR 3,000,000 - 5,000,000
Remote-friendly culture
Learning budget
Performance bonuses
+1
Prismforce Pvt Ltd - AI Engineer - LLM/RAG
Prismforce Pvt Ltd - AI Engineer - LLM/RAG

Prismforce • Maharashtra

On-site
INR 1,800,000 - 3,200,000
Senior AI/ML Engineer
Senior AI/ML Engineer

Keka Technologies Private Limited • Hyderabad

On-site
INR 1,000,000 - 1,500,000
Mentorship from senior architects
Innovation-driven environment
Continuous learning opportunities
AI Engineer (LLM / Generative AI/Ops)
AI Engineer (LLM / Generative AI/Ops)

HCLTech • Dadri

Hybrid
INR 1,500,000 - 4,000,000
AI Developer
AI Developer

Salvo Software • Bengaluru

On-site
INR 1,800,000 - 3,000,000
AI Engineer – LLM & Agentic Systems
AI Engineer – LLM & Agentic Systems

Navaris Digital • Bengaluru

On-site
INR 1,200,000 - 1,800,000