Experience: 3+ Years
Notice Period: Immediate to 15days
Role Overview:
We are looking for a highly driven AI Engineer with hands-on experience in Generative AI and Large Language Models (LLMs). This role offers the opportunity to work on cutting-edge AI solutions, including RAG systems, intelligent agents, and scalable GenAI applications used in real-world business scenarios.
Key Responsibilities:
- Design, develop, and deploy Generative AI solutions leveraging LLMs for real-world use cases.
- Build and optimize RAG (Retrieval-Augmented Generation) pipelines using frameworks like LangChain, LangGraph, and LlamaIndex.
- Develop LLM-powered applications and agents, including multi-step workflows and tool integrations.
- Implement prompt engineering strategies, context management, and response optimization.
- Work on embedding models and vector search systems using tools like Pinecone, FAISS, or similar.
- Develop and maintain scalable backend APIs using FastAPI or Flask for AI applications.
- Collaborate with cross-functional teams to translate business requirements into AI-driven solutions.
- Experiment with fine-tuning, model optimization, and evaluation techniques for LLMs.
- Stay up to date with the latest advancements in Generative AI, agent frameworks, and LLM ecosystems.
Mandatory Skills & Expertise:
- 3+ years of experience in AI/ML, with hands-on exposure to Generative AI projects.
- Strong understanding of LLMs, transformers, and RAG architectures.
- Proficiency in Python and libraries such as PyTorch, TensorFlow, Hugging Face Transformers.
- Experience in building or integrating REST APIs using FastAPI or Flask.
- Familiarity with vector databases like Pinecone, FAISS, or Weaviate.
- Hands-on experience or understanding of prompt engineering, embeddings, and semantic search.
- Strong problem-solving skills and ability to work in a fast-paced environment.