Artificial Intelligence Engineer

iProgrammer Solutions

Pune District

On-site

INR 900,000 - 1,300,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

iProgrammer Solutions in Pune, India is seeking a skilled GenAI engineer to design, build, and scale GenAI-powered applications. You will implement RAG pipelines, vector databases, and LLM orchestration, while integrating with backend systems and DBs.

Strong Python backend skills and practical experience with open-source or cloud LLMs are required. Responsibilities include developing RESTful services with FastAPI, ensuring latency and accuracy, and collaborating with product teams to translate

Qualifications

  • Proven experience building GenAI/LLM-based applications with backend integration.
  • Strong understanding of RAG architecture, embeddings, vector databases and retrieval optimization.
  • Experience with orchestration frameworks like LangChain/LangGraph/LlamaIndex.
  • Familiarity with REST APIs and FastAPI/Django/Flask.
  • Knowledge of LLM guardrails, safety filters, and JSON schema-based outputs.

Responsibilities

  • Design, develop, and maintain GenAI-powered applications (chatbots, support assistants, recommender systems).
  • Build and optimize RAG pipelines with ingestion, chunking, embeddings, and vector search.
  • Implement LLM orchestration and prompt engineering for structured outputs.
  • Develop backend AI services using Python, FastAPI, and microservices.
  • Integrate AI workflows with databases, APIs, CRM/ERP, and payment flows.
  • Create evaluation datasets and failure analyses; collaborate with product and DevOps.

Skills

Python
Backend development
LLM-based apps
LangChain/LangGraph/LlamaIndex
Vector databases
Prompt engineering
REST APIs
FastAPI
Docker
Git
Linux basics
Latency optimization

Education

BE/BTech/MTech/MCA in Computer Science or related

Tools

LangChain
LangGraph
LlamaIndex
FAISS
ChromaDB/Pinecone/Weaviate/Qdrant
Milvus
pgvector
Docker
Git

Job description

Key Responsibilities
  • Design, develop, and maintain GenAI-powered applications such as chatbots, customer support assistants, recommendation assistants, and document/query intelligence systems.
  • Build and optimize RAG pipelines including document ingestion, chunking, embeddings, vector search, retrieval logic, reranking, and response generation.
  • Work on LLM orchestration using frameworks such as LangChain, LangGraph, LlamaIndex, or similar tools.
  • Develop backend AI services using Python, FastAPI, REST APIs, and microservice-based architecture.
  • Implement intent classification, entity extraction, query parsing, and structured JSON output generation for business workflows.
  • Work with open-source and cloud-based LLMs, including local model setup, model serving, and inference optimization.
  • Implement guardrails to prevent hallucination, unsafe responses, competitor comparisons, irrelevant answers, prompt injection, and data leakage.
  • Optimize AI systems for latency, accuracy, reliability, token usage, and scalability.
  • Integrate AI workflows with backend systems, databases, APIs, CRM/ERP systems, payment flows, and business applications.
  • Create evaluation datasets, regression test cases, accuracy reports, and failure analysis for AI responses.
  • Collaborate with product managers, backend developers, DevOps teams, and business stakeholders to convert business requirements into scalable AI solutions.
Required Skills
  • Strong hands-on experience with Python and backend development.
  • Practical experience in building LLM-based applications using OpenAI, Claude, Gemini, Llama, Mistral, or other open-source models.
  • Good understanding of RAG architecture, embeddings, vector databases, semantic search, hybrid search, and retrieval optimization.
  • Experience with frameworks such as LangChain, LangGraph, LlamaIndex, or similar orchestration tools.
  • Knowledge of vector databases such as FAISS, ChromaDB, Pinecone, Weaviate, Qdrant, Milvus, or pgvector.
  • Experience with prompt engineering, system prompts, structured output generation, function calling/tool calling, and JSON schema-based responses.
  • Understanding of LLM guardrails, safety filters, fallback handling, confidence scoring, and hallucination control.
  • Experience with FastAPI / Flask / Django, REST APIs, and backend integration.
  • Understanding of Docker, Git, Linux basics, and deployment workflows.
  • Ability to debug production AI issues related to latency, incorrect responses, token limits, retrieval failure, context mismatch, and model output inconsistency.
Good to Have Skills
  • Experience with vLLM, Ollama, Hugging Face Transformers, TensorRT-LLM, or other model serving frameworks.
  • Knowledge of model quantization, inference optimization, batching, GPU utilization, and token streaming.
  • Experience with AWS, Azure, or GCP for AI/ML deployment.
  • Exposure to telecom, fintech, customer support, billing, recharge, payments, or high-scale consumer applications.
  • Experience in multilingual AI systems, especially Hinglish or Indian language handling.
  • Understanding of ASR/transcription-based input normalization will be a plus.
  • Experience with monitoring tools, logging, prompt/version management, and AI evaluation frameworks.
  • Knowledge of MCP, agentic workflows, tool-based reasoning, or multi-step AI orchestration will be an advantage.
Candidate Profile
  • Experience: 2 to 5 years
  • Education: BE/BTech/MTech/MCA in Computer Science, AI/ML, Data Science, IT, or equivalent practical experience.
  • The candidate should have built at least one real-world AI/GenAI application involving LLMs, backend integration, RAG, or production deployment.
  • The candidate should be able to explain implementation details clearly, not just theoretical AI concepts.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GenAI Engineer (LLM) | Python | RAG | LangChain | AI Integration
GenAI Engineer (LLM) | Python | RAG | LangChain | AI Integration

ProLegion • Jaipur

On-site
INR 1,200,000 - 2,400,000
Artificial Intelligence Developer
Artificial Intelligence Developer

Allegis Group • Hyderabad, Pune District, Bengaluru

Hybrid
INR 1,200,000 - 2,400,000
GenAi Engineer
GenAi Engineer

Habilelabs Private Limited • Gurugram District

On-site
INR 1,200,000 - 1,800,000
AI Engineer
AI Engineer

Easyrewardz Software Services • Gurugram District

On-site
INR 2,500,000 - 4,000,000
Junior AI Engineer
Junior AI Engineer

Ion Enterprise Solutions • Pune District

Hybrid
INR 1,500,000 - 2,500,000
Senior AI Data Engineer
Senior AI Data Engineer

EXL • Maharashtra

On-site
INR 1,200,000 - 1,800,000
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Nxtwave Disruptive Technologies (Hiring for a client) • Pune District, Chennai District, Bengaluru

On-site
INR 2,500,000 - 4,000,000
AI/ML Engineer
AI/ML Engineer

Jade Global Software Pvt LTD • Pune District

On-site
INR 1,200,000 - 1,800,000
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Innovationm • Gurugram District, Bengaluru, Hyderabad

On-site
INR 1,800,000 - 2,800,000
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Questhiring • Gurugram District

On-site
INR 2,500,000 - 4,500,000