Data & Machine Learning Engineer

IDT

Lima Metropolitana

Presencial

PEN 136.193 - 238.338

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Medium is offering a full-time position for a Data/ML Engineer in Lima, Peru. You will design and maintain data pipelines, focusing on AI-driven applications and efficient data processing.

The ideal candidate has over 8 years of experience in data engineering, including MLOps knowledge and proficiency in Python. You will also collaborate with stakeholders to translate business needs into data-driven solutions, all while ensuring robust documentation and compliance.

Formación

  • 8+ years as a Data Engineer with 2+ years focused on MLOps.
  • Deep understanding of vector databases and RAG architectures.
  • Experience with cloud platforms like AWS or Azure Machine Learning.

Responsabilidades

  • Design and maintain scalable data pipelines for model training.
  • Build workflows for extracting unstructured data representations.
  • Create documentation for data processes and model deployment.

Conocimientos

Python for data engineering
MLOps knowledge
Big data technologies (Spark, Hadoop, Kafka)
SQL and PL/SQL programming
Agile methodologies
Version control systems
English communication skills

Herramientas

AWS
Azure Machine Learning
Snowflake
Redshift

Descripción del empleo

This is a full-time opportunity for a data/ML Engineer from LATAM. In‑person verification will be conducted.

IDT is an American telecommunications company founded in 1990 and headquartered in New Jersey. It is an industry leader in prepaid communication and payment services and one of the world’s largest international voice carriers. The company is listed on the NYSE, employs over 1,300 people across 20+ countries, and has revenues in excess of $1.5 billion.

We are looking for a skilled Data/ML Engineer to join our BI team and take an active role in designing, building, and maintaining the end‑to‑end data pipeline, architecture, and design that powers our warehouse, LLM‑driven applications, and AI‑based BI.

Responsibilities
  • Design, develop, and maintain scalable data pipelines to support ingestion, transformation, and delivery into centralized feature stores, model‑training workflows, and real‑time inference services.
  • Build and optimize workflows for extracting, storing, and retrieving semantic representations of unstructured data to enable advanced search and retrieval patterns.
  • Architect and implement lightweight analytics and dashboarding solutions that deliver natural language query experience and AI‑backed insights.
  • Define and execute processes for managing prompt engineering techniques, orchestration flows, and model fine‑tuning routines to power conversational interfaces.
  • Oversee vector data stores and develop efficient indexing methodologies to support retrieval‑augmented generation (RAG) workflows.
  • Partner with data stakeholders to gather requirements for language‑model initiatives and translate them into scalable solutions.
  • Create and maintain comprehensive documentation for all data processes, workflows, and model deployment routines.
  • Stay informed and learn emerging methodologies in data engineering, MLOps, and LLM operations.
Requirements
  • 8+ years of experience as a Data Engineer with 2+ years focused on MLOps.
  • Excellent English communication skills.
  • Effective oral and written communication skills with the BI team and user community.
  • Demonstrated experience in utilizing Python for data engineering tasks, including transformation, advanced data manipulation, and large‑scale data processing.
  • Deep understanding of vector databases and RAG architectures, and how they drive semantic retrieval workflows.
  • Skilled at integrating open‑source LLM frameworks into data engineering workflows for end‑to‑end model training, customization, and scalable inference.
  • Experience with cloud platforms like AWS or Azure Machine Learning for managed LLM deployments.
  • Hands‑on experience with big data technologies including Apache Spark, Hadoop, and Kafka for distributed processing and real‑time data ingestion.
  • Experience designing complex data pipelines extracting data from RDBMS, JSON, API, and flat‑file sources.
  • Demonstrated skills in SQL and PL/SQL programming, with advanced mastery in Business Intelligence and data warehouse methodologies, and hands‑on experience in one or more relational database systems and cloud‑based database services such as Snowflake or Redshift.
  • Understanding of software engineering principles and experience working on Unix/Linux/Windows operating systems, and experience with Agile methodologies.
  • Proficiency in version control systems, with experience in managing code repositories, branching, merging, and collaborating within a distributed development environment.
  • Interest in business operations and comprehensive understanding of how robust BI systems drive corporate profitability by enabling data‑driven decision‑making and strategic insights.
Pluses
  • Experience with vector databases such as DataStax AstraDB, and developing LLM‑powered applications using popular open‑source frameworks like LangChain and LlamaIndex – including prompt engineering, retrieval‑augmented generation (RAG), and orchestration of intelligent workflows.
  • Familiarity with evaluating and integrating open‑source LLM frameworks – such as Hugging Face Transformers or LLaMA‑4 – across end‑to‑end workflows, including fine‑tuning and inference optimization.
  • Knowledge of MLOps tooling and CI/CD pipelines to manage model versioning and automated deployments.

Only accepting applicants from LATAM.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

End-to-End Data & ML Engineer (LATAM) — AI-Driven Pipelines
End-to-End Data & ML Engineer (LATAM) — AI-Driven Pipelines

Medium • San Juan de Lurigancho

Presencial
PEN 136.193 - 238.338
Data Engineer [Zeal]
Data Engineer [Zeal]

Rallyday Partners • Perú

Presencial
PEN 135.000 - 236.000
Data Engineer - Lima
Data Engineer - Lima

Yeah! Global • Lima Metropolitana

Presencial
PEN 133.000 - 234.000
Data Engineer, Azure - Remote, Latin America
Data Engineer, Azure - Remote, Latin America

Bluelight Consulting LLC • Lima Metropolitana

A distancia
PEN 120.000 - 207.000
Competitive salary and performance bonuses
Generous paid-time-off policy
Flexible working hours
+2
Remote Machine Learning Engineer
Remote Machine Learning Engineer

Scopic Software • Lima Metropolitana

Híbrido
Data Engineer, Azure - Remote, Latin America
Data Engineer, Azure - Remote, Latin America

Bluelight Consulting LLC • Piura

A distancia
PEN 103.000 - 207.000
Competitive salary and bonuses
Generous paid-time-off policy
Flexible working hours
+1
Principal AI Engineer I (Python + AWS) – Advanced English) IRC303058
Principal AI Engineer I (Python + AWS) – Advanced English) IRC303058

GlobalLogic • Lima Metropolitana

Presencial
PEN 60.000 - 120.000
Flexible work schedules
English classes
Professional certifications
+1
Remote AI Data Scientist (LATAM) — Part‑Time/Side Gig
Remote AI Data Scientist (LATAM) — Part‑Time/Side Gig

Futureproofing • Lima Metropolitana

Presencial
PEN 136.000 - 205.000
Training cost covered
Flexible schedules
Continuous learning path
Azure Data Engineer - Remote, Latin America
Azure Data Engineer - Remote, Latin America

Bluelight Consulting LLC • Iquitos

Híbrido
PEN 204.000 - 274.000
Competitive salary and bonuses
Generous paid-time-off policy
Flexible working hours
+2
AI Engineer – Agentes Conversacionales & GenAI Workflows
AI Engineer – Agentes Conversacionales & GenAI Workflows

Tekton Labs • Lima Metropolitana

A distancia
PEN 103.000 - 156.000
Proyectos innovadores
Flexibilidad remota
Cultura de aprendizaje continuo