Remote Databricks Data Engineer — Lakehouse & AI Pipelines

AgileEngine

Bogotá ciudad

Presencial

COP 90.000.000 - 150.000.000

Jornada completa

hace 33 horas
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

No envíes un currículum genérico: crea un currículum y una carta de presentación adaptados a este puesto concreto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Flexible remote work
Annual learning budget
Growth opportunities
Competitive compensation
Supportive culture
Remote collaboration with global teams

Descripción de la vacante

AgileEngine is seeking a Middle Data Engineer to modernize a 15-year-old data warehouse into a governed Databricks Lakehouse. Build batch and streaming data pipelines with PySpark and Delta Lake, following a medallion architecture across bronze, silver, and gold layers.

Tools like Claude and Copilot speed up development. You will migrate legacy ETL workloads, optimize Spark jobs, and implement data quality and governance with Unity Catalog, while collaborating with DevOps and analytics teams.

Formación

  • 3+ years of professional data engineering experience.
  • Strong hands-on experience with Databricks, Spark (PySpark), and Delta Lake.
  • Advanced SQL and Python with data modeling skills.
  • Experience with streaming ingestion (Structured Streaming, Auto Loader, Kafka, Event Hubs).
  • Experience with workflow orchestration (Databricks Workflows, Airflow, or Azure Data Factory).
  • Experience with legacy platform migrations or ETL modernization.
  • Familiarity with Unity Catalog, data governance, and PII handling.
  • Experience with dbt or equivalent transformation framework.
  • Upper-intermediate English level.

Responsabilidades

  • Design, build, and operate batch and streaming data pipelines on Databricks using PySpark, Delta Lake, and Databricks Workflows.
  • Model and maintain a medallion architecture serving analytics, reporting, and ML consumers.
  • Migrate legacy ETL and data warehouse workloads onto the Lakehouse with validated data parity and minimal disruption.
  • Use Claude or Copilot to accelerate development and generate tests and docs.
  • Write clean, well-tested Python and SQL with code reviews and documentation.
  • Optimize Spark jobs and Delta tables for performance and cost; manage partitioning and clustering.
  • Implement data quality, lineage, governance via Unity Catalog and automated checks.
  • Debug and resolve production data pipeline incidents; collaborate on observability and security.

Conocimientos

Apache Spark
PySpark
Delta Lake
SQL
Python
Structured Streaming
Databricks Workflows
Data modeling
Data warehouse migration

Herramientas

Databricks
Airflow
Azure Data Factory
Unity Catalog
dbt

Descripción del empleo

AgileEngine is seeking a Middle Data Engineer to modernize a 15-year-old data warehouse into a governed Databricks Lakehouse. Build batch and streaming data pipelines with PySpark and Delta Lake, following a medallion architecture across bronze, silver, and gold layers.

Tools like Claude and Copilot speed up development. You will migrate legacy ETL workloads, optimize Spark jobs, and implement data quality and governance with Unity Catalog, while collaborating with DevOps and analytics teams.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Remote Data Engineer: Databricks Lakehouse & Pipelines
Remote Data Engineer: Databricks Lakehouse & Pipelines

AgileEngine • Metropolitana

Presencial
COP 283.751.000 - 409.862.000
Growth without limits
Competitive compensation
Flexibility
+3
Remote Data Engineer: Databricks Lakehouse & AI Pipelines
Remote Data Engineer: Databricks Lakehouse & AI Pipelines

INGEPSY • Medellín

Presencial
COP 279.616.000 - 466.027.000
Growth opportunities
100% remote work
Annual learning budget
+1
Remote Data Engineer — Databricks Lakehouse & AI Pipelines
Remote Data Engineer — Databricks Lakehouse & AI Pipelines

INGEPSY • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Growth opportunities
Competitive compensation
Remote work
+3
Remote Data Engineer - Databricks Lakehouse & AI Pipelines
Remote Data Engineer - Databricks Lakehouse & AI Pipelines

AgileEngine, LLC. • Bogotá

Presencial
COP 90.000.000 - 150.000.000
Growth budget
Competitive compensation
Remote work
+3
Remote Data Engineer, Databricks Lakehouse
Remote Data Engineer, Databricks Lakehouse

AgileEngine • Perímetro Urbano Barranquilla

Presencial
COP 281.461.000 - 406.555.000
Growth opportunities
Competitive compensation
Remote work
+3
Remote Data Engineer: Databricks Lakehouse & Pipelines
Remote Data Engineer: Databricks Lakehouse & Pipelines

INGEPSY • Perímetro Urbano Barranquilla

Presencial
COP 60.000.000 - 120.000.000
Growth without limits: mentorship,Tech
Competitive compensation with reviews
100% remote with flexible hours
+3
Databricks Data Engineer — Lakehouse, PySpark & AI Tools
Databricks Data Engineer — Lakehouse, PySpark & AI Tools

AgileEngine, LLC. • Medellín

Presencial
COP 289.119.000 - 417.617.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
+3
Remote Data Engineer - Databricks Lakehouse & PySpark
Remote Data Engineer - Databricks Lakehouse & PySpark

AgileEngine • Centrosur

Presencial
COP 100.440.000 - 156.240.000
Growth without limits
Competitive compensation
Flexibility
+3
Data Engineer - Remote, Databricks Lakehouse & Pipelines
Data Engineer - Remote, Databricks Lakehouse & Pipelines

AgileEngine, LLC. • Metropolitana

Presencial
COP 90.000.000 - 140.000.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
Databricks Lakehouse Data Engineer — Remote
Databricks Lakehouse Data Engineer — Remote

AgileEngine, LLC. • Centrosur

Presencial
COP 6.000.000 - 9.000.000
Remote work
Flexible hours
Learning budget
+2