ML + Python , LLM, RAG Engineer( 4 yrs )

Orbion Infotech

India

Remote

INR 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Learning stipend
Mentorship from experienced engineers
Remote-first work environment

Job summary

A leading technology firm is seeking a Machine Learning Engineer — LLM & RAG to design and implement end-to-end RAG pipelines, fine-tune LLMs, and build scalable Python services. Candidates should have 4+ years of experience in ML systems with a strong background in Python, PyTorch, and Retrieval-Augmented Generation. This role offers a remote-first and flexible work environment with competitive compensation and mentorship opportunities.

Qualifications

  • 4+ years of hands‑on experience building ML systems with production deployments.
  • Proven track record of shipping end-to-end ML features.
  • Comfort working remotely across time zones.

Responsibilities

  • Design and implement end-to-end RAG pipelines.
  • Fine-tune and optimize LLMs and embedding models.
  • Build scalable Python services for REST APIs.

Skills

Python
PyTorch
Hugging Face Transformers
LangChain
Retrieval-Augmented Generation
FAISS
Docker
REST APIs

Education

4+ years of experience in ML systems

Tools

Docker
AWS
Azure ML
Milvus
Pinecone

Job description

Base pay range

About The Opportunity

Industry: Enterprise Generative AI & Natural Language Processing (NLP). We build LLM-driven search, knowledge augmentation, and intelligent automation solutions for B2B SaaS and enterprise customers. The team focuses on production-grade Retrieval-Augmented Generation (RAG), embedding pipelines, and low-latency inference services that power customer-facing products and internal automation.

Standardized Title: Machine Learning Engineer — LLM & RAG (best-performing title for this search)

Role & Responsibilities
  • Design and implement end-to-end RAG pipelines: document ingestion, embedding generation, vector indexing, retrieval, and prompt orchestration for production LLM applications.
  • Fine-tune, evaluate, and optimize LLMs and embedding models to meet task-specific accuracy, latency, and cost targets.
  • Build scalable Python services to expose inference and retrieval through secure REST APIs and microservices.
  • Integrate vector databases and search (FAISS/Pinecone/Milvus) and implement efficient nearest-neighbor search, caching, and sharding strategies.
  • Containerize and productionize models/services using Docker and orchestration best practices; collaborate on CI/CD, monitoring, and observability for ML workloads.
  • Work cross-functionally with Product, Data Engineering, and DevOps to define KPIs, run A/B tests, and iterate on model quality and user experience.
Skills & Qualifications
Must-Have (Technical Skills)
  • Python
  • PyTorch
  • Hugging Face Transformers
  • LangChain
  • Retrieval-Augmented Generation
  • FAISS
  • Docker
  • REST APIs
Preferred
  • Pinecone
  • Milvus
  • AWS (SageMaker / EC2) or Azure ML
Qualifications
  • 4+ years of hands‑on experience building ML systems with production deployments in LLM/RAG or NLP applications.
  • Proven track record of shipping end-to-end ML features: data ingestion, training/fine‑tuning, inference, and monitoring.
  • Comfort working remotely across time zones and collaborating asynchronously with engineering and product teams in India.
Benefits & Culture Highlights
  • Remote‑first, flexible‑work environment with emphasis on ownership and learning.
  • Opportunities to work on cutting‑edge LLM projects and influence product direction.
  • Competitive compensation, learning stipend, and mentorship from experienced ML engineers.

This role is optimized for engineers who combine strong software engineering discipline with deep practical experience in LLMs, vector search, and production ML. If you enjoy turning research‑grade models into reliable, scalable services, this is an excellent opportunity to make measurable product impact.

Skills: python, system design, rag, llm

Seniority level

Mid-Senior level

Employment type

Full-time

Job function

Other

Industries

IT Services and IT Consulting

Location: Vishakhapatnam, Andhra Pradesh, India

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Engineer
LLM Engineer

Zorba Consulting • Pune District

On-site
INR 1,500,000 - 2,100,000
GPU access on-site
Cross-functional product teams
Ownership and rapid iteration
LLM Engineer
LLM Engineer

Zorba Consulting • Hyderabad

On-site
INR 2,500,000 - 4,500,000
On-site role with GPU access
Cross-functional product teams
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Amerisource Solutions • Hyderabad

On-site
INR 1,800,000 - 3,200,000
Senior AI / LLM Engineer
Senior AI / LLM Engineer

Thinkscoop Technologies • Bengaluru

Hybrid
INR 3,000,000 - 5,000,000
Remote-friendly culture
Learning budget
Performance bonuses
+1
Principal Engineer
Principal Engineer

LE400 Automation Anywhere Software Pvt. Ltd. • Bengaluru

On-site
INR 2,800,000 - 4,200,000
Senior AI/ML Engineer
Senior AI/ML Engineer

Ciklum India • Maharashtra

On-site
INR 2,000,000 - 3,000,000
Company-paid medical insurance
Mental health support
Financial and legal consultations
+1
AI/ML Engineer
AI/ML Engineer

TecOrb Technologies - We Believe in Challenges • Dadri

On-site
INR 600,000 - 1,000,000
Competitive compensation
Performance-based incentives
Collaborative startup culture
Lead AI/ML Engineer
Lead AI/ML Engineer

Relanto • Bengaluru

On-site
INR 1,500,000 - 2,000,000
Senior AI/ML Engineer
Senior AI/ML Engineer

Keka Technologies Private Limited • Hyderabad

On-site
INR 1,000,000 - 1,500,000
Mentorship from senior architects
Innovation-driven environment
Continuous learning opportunities
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Aventra.AI • Gurugram District

On-site
INR 450,000 - 900,000