Machine Learning Engineer - AI

Jobgether

India

On-site

INR 1,500,000 - 2,800,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Edge deployment exposure
MLOps pipelines
Hugging Face tech
Team collaboration

Job summary

Jobgether is seeking a Machine Learning Engineer - AI based in India. The role focuses on fine-tuning small language models, edge deployment, and production-ready MLOps across edge, mobile, and local environments.

You will work with Hugging Face, TRL, LoRA, QLoRA, and PEFT, building end-to-end pipelines, monitoring latency and accuracy, and collaborating with cross-functional teams to deliver robust AI solutions.

Qualifications

  • Hands-on experience developing, training, and fine-tuning Small Language Models or transformer-based models.
  • Strong experience with Hugging Face and adaptation techniques (LoRA/QLoRA/PEFT).
  • Experience deploying ML models to edge/mobile/local environments with latency constraints.

Responsibilities

  • Fine-tune and train SLMs using Hugging Face, TRL, LoRA, QLoRA, and PEFT.
  • Experiment with architectures, datasets and strategies to improve performance.
  • Optimize models for efficient inference (quantization, pruning, distillation).
  • Deploy lightweight models to edge/mobile/local servers and constrained devices.
  • Design end-to-end MLOps pipelines for data, training, deployment, and monitoring.
  • Build repeatable workflows to support model development and production deployment.
  • Monitor accuracy, latency, and hardware utilization in production.
  • Develop benchmarking frameworks and evaluate model performance.
  • Collaborate with cross-functional teams to integrate models into products.
  • Advance ML engineering best practices in experimentation, observability, and CI/CD.
  • Explore new techniques and tooling for efficient AI inference and edge deployment.

Skills

Hugging Face
LoRA
QLoRA
PEFT
MLOps
Edge deployment
Small Language Models
Model quantization
ONNX export

Tools

ONNX export
MLOps tooling

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Machine Learning Engineer - AI based in India.

This is an engineering-focused opportunity for a machine learning professional working at the intersection of small language models, model optimization, and production AI. You will fine-tune and optimize lightweight models designed for efficient inference across edge, mobile, and local environments. The role spans the full machine learning lifecycle, from data ingestion and experimentation through deployment, monitoring, and continuous improvement. You will work with modern frameworks and techniques such as Hugging Face, TRL, LoRA, QLoRA, and PEFT to build efficient AI solutions. A strong focus is placed on performance, with strict attention to latency, model accuracy, and hardware utilization. You will also contribute to scalable MLOps practices that make AI models reliable and production-ready.

Accountabilities:

  • Fine-tune and train Small Language Models (SLMs) using Hugging Face, TRL, and parameter-efficient adaptation techniques such as LoRA, QLoRA, and PEFT.
  • Experiment with model architectures, training strategies, and datasets to improve model quality and task-specific performance.
  • Optimize models for efficient inference using techniques including quantization, pruning, knowledge distillation, and other model compression approaches.
  • Prepare and deploy lightweight AI models to edge devices, mobile environments, local servers, and other resource-constrained platforms.
  • Design and implement end-to-end MLOps pipelines covering data ingestion, preprocessing, experimentation, model training, validation, packaging, deployment, and monitoring.
  • Build reliable and repeatable workflows that support efficient model development and production deployment.
  • Monitor deployed models for accuracy, latency, resource consumption, and CPU/GPU utilization.
  • Develop and maintain model benchmarking frameworks and custom evaluation suites to measure model quality and performance.
  • Analyze production performance and identify opportunities to improve model efficiency, reliability, and scalability.
  • Work with cross-functional engineering teams to integrate machine learning models into real-world products and environments.
  • Contribute to ML engineering best practices around experimentation, versioning, deployment, observability, and continuous improvement.
  • Explore emerging techniques and tooling for efficient AI inference, edge deployment, and production machine learning.
Requirements:
  • Hands-on experience developing, training, and fine-tuning Small Language Models or other transformer-based models.
  • Strong practical knowledge of Hugging Face and modern model adaptation techniques, including LoRA, QLoRA, and PEFT.
  • Experience optimizing machine learning models for efficient inference through quantization, pruning, knowledge distillation, or similar techniques.
  • Experience deploying machine learning models to edge devices, mobile platforms, local servers, or other environments with constrained compute and strict latency requirements.
  • Strong understanding of end-to-end MLOps practices, from data ingestion and model experimentation through deployment and production monitoring.
  • Experience monitoring model accuracy, inference latency, and CPU/GPU or other hardware utilization in production.
  • Ability to develop meaningful model evaluation and benchmarking frameworks and use data-driven results to improve model performance.
  • Strong software engineering, debugging, analytical, and problem-solving skills.
  • Ability to work effectively in a collaborative, fast-moving environment and communicate technical concepts clearly.
  • Strong ownership mindset and willingness to continuously learn new machine learning technologies and deployment techniques.
  • Experience with ONNX export and cross-platform inference is a plus.
  • Experience deploying AI solutions to edge or mobile environments is preferred.
  • Familiarity with MLOps tooling for experiment tracking, model registries, and ML-focused CI/CD pipelines is advantageous.
Benefits:
  • Opportunity to work on modern AI and machine learning technologies with a strong focus on Small Language Models.
  • Hands-on exposure to model fine-tuning, optimization, compression, and efficient inference.
  • Opportunity to build production-grade MLOps pipelines spanning the full machine learning lifecycle.
  • Experience working with edge, mobile, and local AI deployment environments.
  • Exposure to technologies including Hugging Face, TRL, LoRA, QLoRA, PEFT, and modern MLOps tooling.
  • Opportunity to contribute to scalable AI solutions designed for real-world production environments.
  • Collaborative environment with opportunities to work alongside experienced engineering and technology professionals.
  • Strong focus on continuous learning, experimentation, and adoption of emerging AI technologies.
  • Inclusive workplace culture that values diverse perspectives, collaboration, and individual contributions.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Engineer_6+yrs
Machine Learning Engineer_6+yrs

Zorba AI • India

Hybrid
INR 1,500,000 - 2,100,000
Machine Learning Engineer – MLOps & AI
Machine Learning Engineer – MLOps & AI

CNIM MARTIN Pvt. Ltd • Chennai District

On-site
INR 1,200,000 - 1,800,000
Senior AI/ML Engineer
Senior AI/ML Engineer

Unifyed • Gurgaon

On-site
INR 1,500,000 - 2,500,000
Machine Learning Engineer
Machine Learning Engineer

SWITS DIGITAL Private Limited • Chennai District

On-site
INR 3,000,000 - 6,000,000
ML Engineer
ML Engineer

TAAS Partners • Bengaluru

On-site
INR 1,800,000 - 2,800,000
Salary + equity options
Professional growth opportunities
Benefits package
Senior ML Backend Engineer
Senior ML Backend Engineer

Jobgether • India

On-site
INR 3,500,000 - 6,500,000
Advanced ML projects
Geospatial analytics exposure
Cloud-native infra
+2
Software Engineer - AI Software & Platform
Software Engineer - AI Software & Platform

United States Digital Space LLC • Karnataka

On-site
INR 2,500,000 - 4,500,000
ML Engineer
ML Engineer

Cubet Techno Labs • Ernakulam

On-site
INR 700,000 - 1,200,000
Innovative Environment
Career Growth
Collaborative Culture
+1
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Nifty Coders Pvt. Ltd. • Nalanda

Hybrid
INR 2,500,000 - 4,000,000
Competitive compensation plans
Two annual bonuses
Paid Maternity Leave (4 months) and P.
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Recro • Hyderabad

On-site
INR 900,000 - 1,800,000