Senior Data Engineer (with AI/ML experience) India

IDT

Pune District

On-site

INR 4,200,000 - 6,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Remote work opportunity
Annual performance review
Career growth opportunities
Company-supported English classes
Referral program

Job summary

IDT is seeking a Senior Data Engineer (with AI/ML exposure) to join the BI team and own end-to-end data pipelines powering the warehouse, LLM-driven apps, and AI-based BI. The role emphasizes scalable architecture, vector storage, and RAG-informed retrieval, with a strong emphasis on English communication and variable work locations.

We offer remote work options and a competitive salary with opportunities for growth, training, and cross-functional collaboration across global teams.

Qualifications

  • Senior Data Engineer with 6-7+ years of experience in building scalable data infrastructure.
  • - 1-2 years hands-on exposure to AI/ML workflows.
  • - Excellent English communication skills.
  • - Experience with big data technologies (Spark, Hadoop, Kafka).
  • - Strong SQL/PLSQL and data warehousing knowledge (Snowflake/Redshift).
  • - Proficiency in Python for data engineering tasks.
  • - Familiarity with vector databases and RAG architectures.
  • - Experience integrating open-source LLMs into data pipelines.
  • - Cloud experience (AWS or Azure ML).
  • - Agile methodologies and version control (Git).

Responsibilities

  • Design, develop, and maintain scalable data pipelines powering warehouses, feature stores, model-training workflows, and real-time services.
  • Design and optimize ETL/ELT pipelines and data structures in cloud data warehouses (Snowflake/Redshift).
  • Build and optimize workflows for semantic representations of unstructured data for advanced search and retrieval.
  • Develop lightweight analytics and dashboards with AI-backed insights.
  • Define processes for prompt engineering, orchestration, and model fine-tuning for conversational interfaces.
  • Oversee vector data stores and indexing for retrieval-augmented generation workflows.
  • Collaborate with data stakeholders to translate requirements into scalable solutions.
  • Document data processes, workflows, and deployment routines.
  • Stay updated on emerging data engineering, MLOps, and LLM operations.

Skills

Spark
Hadoop
Kafka
SQL
Python
Vector databases
RAG
LLM frameworks
MLOps
AWS
Azure
Unix/Linux
Git

Tools

Snowflake
Redshift

Job description

IDT (www.idt.net) is a communications and financial services company founded in 1990 and headquartered in New Jersey, US. Today it is an industry leader in prepaid communication and payment services and one of the world’s largest international voice carriers. We are listed on the NYSE, employ over 1800 people across 20+ countries, and have revenues in excess of $1.5 billion.

We are looking for a skilled Senior Data Engineer (with AI/ML exposure) to join our BI team and take an active role in designing, building, and maintaining the end-to-end data pipeline, architecture and design that powers our warehouse, LLM-driven applications, and AI-based BI. If you're looking for a company that will give you the maximum flexibility in choosing a location to work, this opportunity is for you!

The interview process will be conducted in English.

Responsibilities:
  • Design, develop, and maintain scalable data pipelines to support ingestion, transformation, and delivery into centralized feature stores, model-training workflows, and real-time inference services.
  • Design, optimize, and maintain robust ETL/ELT pipelines and data structures within our cloud data warehouse (Snowflake/Redshift) to support core Business Intelligence.
  • Build and optimize workflows for extracting, storing, and retrieving semantic representations of unstructured data to enable advanced search and retrieval patterns.
  • Architect and implement lightweight analytics and dashboarding solutions that deliver natural language query experience and AI-backed insights.
  • Define and execute processes for managing prompt engineering techniques, orchestration flows, and model fine-tuning routines to power conversational interfaces.
  • Oversee vector data stores and develop efficient indexing methodologies to support retrieval-augmented generation (RAG) workflows.
  • Partner with data stakeholders to gather requirements for language-model initiatives and translate into scalable solutions.
  • Create and maintain comprehensive documentation for all data processes, workflows and model deployment routines.
  • Should be willing to stay informed and learn emerging methodologies in data engineering, MLOps and LLM operations.
Requirements:
  • Senior Data Engineer with 6-7+ years of experience in building scalable data infrastructure and 1-2 years of hands-on exposure to AI/ML workflows.
  • Excellent English communication skills.
  • Hands-on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real-time data ingestion.
  • Experience designing complex data pipelines extracting data from RDBMS, JSON, API and Flat file sources.
  • Demonstrated skills in SQL and PLSQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, along with hands-on experience in one or more relational database systems and cloud-based database services such as Snowflake/Redshift.
  • Effective oral and written communication skills with BI team and user community.
  • Demonstrated experience in utilizing python for data engineering tasks, including transformation, advanced data manipulation, and large-scale data processing.
  • Deep understanding of vector databases and RAG architectures, and how they drive semantic retrieval workflows.
  • Skilled at integrating open-source LLM frameworks into data engineering workflows for end-to-end model training, customization, and scalable inference.
  • Experience with cloud platforms like AWS or Azure Machine Learning for managed LLM deployments.
  • Understanding of software engineering principles and skills working on Unix/Linux/Windows Operating systems, and experience with Agile methodologies.
  • Proficiency in version control systems, with experience in managing code repositories, branching, merging, and collaborating within a distributed development environment.
  • Interest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data-driven decision-making and strategic insights.
Pluses:
  • Experience with vector databases such as DataStax AstraDB, and developing LLM-powered applications using popular open source frameworks like LangChain and LlamaIndex–including prompt engineering, retrieval-augmented generation (RAG), and orchestration of intelligent workflows.
  • Familiarity with evaluating and integrating open-source LLM frameworks–such as Hugging Face Transformers/LLaMA-4 across end-to-end workflows, including fine-tuning and inference optimization.
  • Knowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments.
What we offer:
  • Remote work opportunity!
  • B2B Employment ($, gross) or full-time employment option.
  • Stable job with long-term growth perspective.
  • Competitive salary with annual performance review.
  • Really good hardware.
  • An exciting and challenging job with talented people around.
  • Continuous learning and career growth opportunities.
  • Compensation for professional training, seminars, and conferences.
  • Referral program – get rewarded for helping us grow the team with talented people.
  • Company-supported English classes to enhance your professional growth.

Only accepting applicants from INDIA

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer (with AI/ML experience) India
Senior Data Engineer (with AI/ML experience) India

IDT • Bengaluru

On-site
INR 1,400,000 - 2,100,000
Remote work opportunity
Competitive salary
Long-term growth
Senior Data Engineer (with AI/ML experience) India
Senior Data Engineer (with AI/ML experience) India

IDT Corporation • Bengaluru

Remote
INR 6,433,000 - 9,192,000
Competitive salary
Stability and growth opportunities
Referral program
+2
(Senior) AI Engineer - Data & AI Organisation (all genders) 1
(Senior) AI Engineer - Data & AI Organisation (all genders) 1

Merck Group • Bengaluru

On-site
INR 1,400,000 - 2,000,000
ML Engineer || Contract Job || 8-15 Years Experience
ML Engineer || Contract Job || 8-15 Years Experience

People Prime Worldwide • Hyderabad

Remote
Senior Data Engineer
Senior Data Engineer

Hiring Hub • Sahibzada Ajit Singh Nagar

On-site
INR 2,000,000 - 3,500,000
Senior Data Engineer
Senior Data Engineer

ACS International India Pvt. Ltd. (ACSII) • Maharashtra

On-site
INR 1,200,000 - 1,800,000
Senior Data Engineer
Senior Data Engineer

Impact Makers • Hyderabad

On-site
INR 1,200,000 - 2,400,000
AI and Data Engineering Tech Lead
AI and Data Engineering Tech Lead

Carelon Global Solutions • Bengaluru

On-site
INR 4,500,000 - 7,500,000
Senior Data / ML Engineer
Senior Data / ML Engineer

Web Spiders Group • Kolkata District

On-site
INR 1,800,000 - 2,800,000
Senior AI/ML Engineer
Senior AI/ML Engineer

Unifyed • Gurgaon

On-site
INR 1,500,000 - 2,500,000