SDE II - ML/AI Engineer

Auriga IT

Jaipur

On-site

INR 1,200,000 - 2,800,000

Full time

39 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Auriga IT is seeking a hands-on GenAI & Computer Vision Engineer with 3–5 years of experience to deliver production‑grade AI solutions. You will own end‑to‑end model development—from research and fine‑tuning through deployment, monitoring, and iteration, in Jaipur.

You will tackle domain challenges like LLM hallucinations, vector search scalability, and real‑time inference constraints while collaborating with cross‑functional teams to build scalable pipelines using vLLM, LangChain, PyTorch, and

Qualifications

  • Bachelor’s/Master’s in CS/EE/AI/ML or related field.
  • 3–5 years shipping generative and vision-based AI models in production.
  • Strong problem-solving mindset; ability to debug LLM drift and model degradation.
  • Excellent verbal and written communication skills.

Responsibilities

  • Generative AI & LLM engineering: fine-tune and evaluate LLMs; deploy high-throughput inference pipelines; design retrieval-augmented generation workflows; build scalable inference APIs.
  • Computer Vision development: develop and optimize CV models for detection, segmentation, classification, and tracking; real-time pipelines; edge/cloud deployment optimization.
  • MLOps & Deployment: containerize with Docker; orchestrate with Kubernetes; implement CI/CD for model/versioning; monitor performance and costs.
  • Cross-functional collaboration: define SLAs for latency/throughput; promote best practices; mentor engineers on reproducible research and end-to-end AI delivery.

Skills

Python
C++
Go
Communication
Problem-solving

Education

Bachelor’s or Master’s in Computer Science, Electrical Engineering, AI/ML, or related field

Tools

Hugging Face Transformers
Ollama
vLLM
LLaMA
LangChain
LangGraph
Pinecone
Weaviate
Milvus
Triton Inference Server
FastAPI
Flask
PyTorch
TensorFlow
OpenCV
NVIDIA DeepStream
TensorRT
ONNX Runtime
Torch-TensorRT
Docker
Kubernetes
KServe
SageMaker
MLflow
DVC
Prometheus
Grafana

Job description

Job Summary

We’re seeking a hands‑on GenAI & Computer Vision Engineer with 3–5 years of experience delivering production‑grade AI solutions. You must be fluent in the core libraries, tools, and cloud services listed below, and able to own end‑to‑end model development—from research and fine‑tuning through deployment, monitoring, and iteration. In this role, you’ll tackle domain‑specific challenges like LLM hallucinations, vector search scalability, real‑time inference constraints, and concept drift in vision models.

Key Responsibilities
Generative AI & LLM Engineering
  • Fine‑tune and evaluate LLMs (Hugging Face Transformers, Ollama, LLaMA) for specialized tasks
  • Deploy high‑throughput inference pipelines using vLLM or Triton Inference Server
  • Design agent‑based workflows with LangChain or LangGraph, integrating vector databases (Pinecone, Weaviate) for retrieval‑augmented generation
  • Build scalable inference APIs with FastAPI or Flask, managing batching, concurrency, and rate‑limiting
Computer Vision Development
  • Develop and optimize CV models (YOLOv8, Mask R‑CNN, ResNet, EfficientNet, ByteTrack) for detection, segmentation, classification, and tracking
  • Implement real‑time pipelines using NVIDIA DeepStream or OpenCV (cv2); optimize with TensorRT or ONNX Runtime for edge and cloud deployments
  • Handle data challenges—augmentation, domain adaptation, semi‑supervised learning—and mitigate model drift in production
MLOps & Deployment
  • Containerize models and services with Docker; orchestrate with Kubernetes (KServe) or AWS SageMaker Pipelines
  • Implement CI/CD for model/version management (MLflow, DVC), automated testing, and performance monitoring (Prometheus + Grafana)
  • Manage scalability and cost by leveraging cloud autoscaling on AWS (EC2/EKS), GCP (Vertex AI), or Azure ML (AKS)
Cross‑Functional Collaboration
  • Define SLAs for latency, accuracy, and throughput alongside product and DevOps teams
  • Evangelize best practices in prompt engineering, model governance, data privacy, and interpretability
  • Mentor junior engineers on reproducible research, code reviews, and end‑to‑end AI delivery
Required Qualifications

You must be proficient in at least one tool from each category below:

LLM Frameworks & Tooling:
  • Hugging Face Transformers, Ollama, vLLM, or LLaMA
Agent & Retrieval Tools:
  • LangChain or LangGraph; RAG with Pinecone, Weaviate, or Milvus
Inference Serving:
  • Triton Inference Server; FastAPI or Flask
Computer Vision Frameworks & Libraries:
  • PyTorch or TensorFlow; OpenCV (cv2) or NVIDIA DeepStream
Model Optimization:
  • TensorRT; ONNX Runtime; Torch‑TensorRT
MLOps & Versioning:
  • Docker and Kubernetes (KServe, SageMaker); MLflow or DVC
Monitoring & Observability:
  • Prometheus; Grafana
Cloud Platforms:
  • AWS (SageMaker, EC2/EKS) or GCP (Vertex AI, AI Platform) or Azure ML (AKS, ML Studio)
Programming Languages:
  • Python (required); C++ or Go (preferred)
Additionally
  • Bachelor’s or Master’s in Computer Science, Electrical Engineering, AI/ML, or a related field
  • 3–5 years of professional experience shipping both generative and vision‑based AI models in production
  • Strong problem‑solving mindset; ability to debug issues like LLM drift, vector index staleness, and model degradation
  • Excellent verbal and written communication skills
Typical Domain Challenges You’ll Solve
  • LLM Hallucination & Safety: Implement grounding, filtering, and classifier layers to reduce false or unsafe outputs
  • Vector DB Scaling: Maintain low‑latency, high‑throughput similarity search as embeddings grow to millions
  • Inference Latency: Balance batch sizing and concurrency to meet real‑time SLAs on cloud and edge hardware
  • Concept & Data Drift: Automate drift detection and retraining triggers in vision and language pipelines
  • Multi‑Modal Coordination: Seamlessly orchestrate data flow between vision models and LLM agents in complex workflows
About Company

Hi there! We are Auriga IT.

We power businesses across the globe through digital experiences, data and insights. From the apps we design to the platforms we engineer, we're driven by an ambition to create world‑class digital solutions and make an impact. Our team has been part of building the solutions for the likes of Zomato, Yes Bank, Tata Motors, Amazon, Snapdeal, Ola, Practo, Vodafone, Meesho, Volkswagen, Droom and many more.

We are a group of people who just could not leave our college‑life behind and the inception of Auriga was solely based on a desire to keep working together with friends and enjoying the extended college life.

Who Has not Dreamt of Working with Friends for a Lifetime

Our Website - https://aurigait.com/

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SDE II - ML AI Engineer
SDE II - ML AI Engineer

Auriga IT • Jaipur

On-site
INR 1,800,000 - 3,000,000
SDE II - ML/AI Engineer
SDE II - ML/AI Engineer

Aurigait • Jaipur

On-site
INR 1,200,000 - 1,800,000
AI Engineer - Machine Learning Models
AI Engineer - Machine Learning Models

VAYUZ Technologies • Bengaluru Urban

On-site
INR 4,000,000 - 7,000,000
AI Tech Lead / LLM Architect / Generative AI Lead
AI Tech Lead / LLM Architect / Generative AI Lead

Avaali Solutions • Bengaluru

On-site
INR 3,000,000 - 7,000,000
Senior AI Engineer
Senior AI Engineer

Zycus Inc. • India

On-site
INR 1,800,000 - 3,200,000
AI/ML Engineer
AI/ML Engineer

Lifesight • Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Health insurance
Daily breakfast
Weekday lunches
+3
AI / ML Engineer
AI / ML Engineer

NITYO • Pune District

On-site
INR 1,400,000 - 2,200,000
GenAI Trainee
GenAI Trainee

Top Gen AI Jobs • Kurla

On-site
INR 573,000 - 1,032,000
Technical Architect- Remote, India
Technical Architect- Remote, India

LeewayHertz Technologies Pvt. Ltd. • India

Remote
INR 4,000,000 - 9,000,000
Remote work in India
AI Engineer
AI Engineer

Teleradiology Solutions • Bengaluru

On-site
INR 1,200,000 - 2,000,000