ETL Developer

Merkleinnovation

Asia

Presencial

PEN 60.000 - 90.000

Jornada completa

Hace 2 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Una candidatura completa en un minuto — currículum adaptado y carta de presentación, listos para enviar.

Supera los filtros ATS

Descripción de la vacante

Merkleinnovation is seeking a skilled ETL Development Engineer with hands-on experience building high-throughput data pipelines using PySpark. You will design, develop, and optimize scalable data warehouse models and Spark-based ETL workflows, collaborating with cross-functional teams to ensure data availability and reliable delivery.

The role emphasizes performance tuning of Spark jobs, DAG orchestration with Airflow, and adherence to coding standards, documentation, and testing.

Formación

  • 3+ years PySpark offline ETL development experience.
  • Strong Spark Core/SQL tuning and Airflow orchestration.
  • Data warehouse modeling with dimensional modeling experience.
  • Familiarity with cloud environment workflows and Telco data processing.
  • Independent delivery with strong coding standards.
  • Experience migrating from legacy ETL systems (Ab Initio) to PySpark.
  • Data mapping and large-scale migrations in Telco/Banking sectors.

Responsabilidades

  • ETL Pipeline Development: design, build, and maintain high-throughput offline batch ETL pipelines using Spark.
  • Performance Tuning: optimize Spark jobs, memory, and query execution for performance and cost efficiency.
  • Data Warehouse Modeling: architect scalable dimensional models for analytical workflows.
  • Pipeline Orchestration: manage DAG workflows using Airflow with monitoring and alerts.
  • Code Quality & Best Practices: ensure clean code, documentation, reviews, and unit testing.
  • Migration & Modernization: participate in or lead ETL re-engineering and platform migrations.
  • Cross-functional Collaboration: read docs and communicate with international teams in English.

Conocimientos

PySpark
ETL development
Performance tuning
Data warehouse modeling
Pipeline orchestration
Code quality
Migration & modernization
Cross-functional collaboration
English communication

Educación

Bachelor's degree in Computer Science or related field

Herramientas

Apache Spark
Airflow

Descripción del empleo

We are looking for a skilled and self-motivated ETL Development Engineer with strong hands-on experience in high-throughput data pipelines and offline batch processing. In this role, you will design, develop, and optimize scalable data warehouse models and Spark-based ETL workflows. You will work closely with cross-functional teams to ensure high data availability, seamless task orchestration, and reliable data architecture delivery.

Responsibilities
  • ETL Pipeline Development: Design, build, and maintain high-throughput offline batch ETL pipelines utilizing Apache Spark (Core/SQL).
  • Performance Tuning: Troubleshoot, benchmark, and optimize Spark jobs, memory allocation, and query execution for high performance and cost efficiency.
  • Data Warehouse Modeling: Architect and implement scalable data warehouse dimensional models (Kimball/Inmon methodologies, ODS/DWH/ADS layering) supporting analytical workflows.
  • Data Warehouse Modeling: Architect and implement scalable data warehouse dimensional models (Kimball/Inmon methodologies, ODS/DWH/ADS layering) supporting analytical workflows.
  • Pipeline Orchestration: Manage DAG workflows, incremental processing, failure retry logic, and real-time monitoring/alerting using orchestration platforms such as Apache Airflow.
  • Code Quality & Best Practices: Maintain high software engineering standards through clean code writing, robust documentation, code reviews, and automated unit testing.
  • Migration & Modernization: Assist in or lead large-scale ETL re-engineering, platform migrations, and optimization initiatives across legacy and modern data stacks.
  • Cross-functional Collaboration: Read technical documentation, communicate seamlessly with international tech teams in English, and deliver features independently from requirements to production.
Requirements
  • Bachelor's degree or above in a computer-related major
  • Experience: 3+ years in PySpark offline ETL development
  • Technical Skills: Spark Core/SQL tuning, Airflow orchestration, and data warehouse modeling.
  • Domain & Infra: Experience or familiarity with cloud environment workflows and telco data processing challenges.
  • Delivery: Strong coding standards and proven independent delivery capability.
  • Language: Working proficiency in English (technical documentation, written, and spoken).
  • Preferred: Prior experience in large-scale ETL migration or reengineering in telco/cloud environments.
  • Proven hands-on experience in executing end-to-end data and logic-level migrations from legacy ETL systems (specifically Ab Initio) to modern frameworks using PySpark
  • Strong background in complex data mapping, table design, and script rewriting, preferably within large-scale Data Migration projects in the Telecommunications or Banking sectors
Let's Build Something Meaningful Together

Helping businesses transform ideas into impactful digital solutions.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

ETL Developer (Legacy ETL Expert)
ETL Developer (Legacy ETL Expert)

Merkleinnovation • Asia

Presencial
PEN 204.000 - 306.000
Azure Data Engineer - Remote, Latin America
Azure Data Engineer - Remote, Latin America

Bluelight Consulting LLC • Huancayo

Híbrido
PEN 204.000 - 274.000
Competitive salary
Generous paid-time-off policy
Flexible working hours
+2
Azure Data Engineer - Remote, Latin America
Azure Data Engineer - Remote, Latin America

Bluelight Consulting LLC • Iquitos

Híbrido
PEN 204.000 - 274.000
Competitive salary and bonuses
Generous paid-time-off policy
Flexible working hours
+2
Data Engineer, Azure - Remote, Latin America
Data Engineer, Azure - Remote, Latin America

Bluelight Consulting LLC • Piura

A distancia
PEN 103.000 - 207.000
Competitive salary and bonuses
Generous paid-time-off policy
Flexible working hours
+1
Data Engineer, Azure - Remote, Latin America
Data Engineer, Azure - Remote, Latin America

Bluelight Consulting LLC • Lima Metropolitana

A distancia
PEN 120.000 - 207.000
Competitive salary and performance bonuses
Generous paid-time-off policy
Flexible working hours
+2
DATA ENGINEER SUPERVISOR
DATA ENGINEER SUPERVISOR

PT Tempo Scan Pacific Tbk • Asia

Presencial
PEN 90.000 - 150.000
Junior ETL Developer
Junior ETL Developer

PT Astra Graphia Information Technology • Asia

Presencial
PEN 60.000 - 90.000
Data Engineer [Zeal]
Data Engineer [Zeal]

Rallyday Partners • Perú

Presencial
PEN 135.000 - 236.000
Senior Data Engineer
Senior Data Engineer

Athenaworks • Lima Metropolitana

Presencial
PEN 202.000 - 369.000
Payment in USD
Flexible work schedule
Non-Working Pay Days Policy
+1
Data Engineer - Lima
Data Engineer - Lima

Yeah! Global • Lima Metropolitana

Presencial
PEN 133.377 - 233.410