Mid Data Engineer - Databricks Lakehouse (Remote)

INGEPSY

Cartagena de Indias

Presencial

COP 279.616.000 - 372.821.000

Jornada completa

Hace 6 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Transforma esta oferta en una entrevista: un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Growth opportunities
Competitive compensation
100% remote with flexible hours
Modern tech stack
Collaborative culture
Well-being & support

Descripción de la vacante

AgileEngine is seeking a Middle Data Engineer to modernize a 15-year data warehouse into a governed Databricks Lakehouse. You will build batch and streaming pipelines with PySpark and Delta Lake, using a medallion architecture for analytics, reporting and ML consumption.

Responsibilities include migrating legacy workloads, implementing data quality and governance, and collaborating with DevOps and analytics teams. Strong SQL/Python, Spark, and cloud architecture experience are required.

Formación

  • 3+ years of professional experience in data engineering.
  • Hands-on with Databricks, Spark (PySpark), and Delta Lake.
  • Advanced SQL and Python, data modeling skills.
  • Experience with Structured Streaming and Kafka.
  • Experience with Databricks Workflows, Airflow, or Azure Data Factory.
  • Experience with legacy migrations and data hygiene.
  • Familiarity with Unity Catalog and data governance.
  • Knowledge of dbt or similar transformation framework.
  • Secure coding and production data platforms at scale.
  • Upper-intermediate English level.

Responsabilidades

  • Design, build, and operate batch and streaming pipelines on Databricks.
  • Model a medallion architecture (bronze/silver/gold) for analytics and ML.
  • Migrate legacy ETL workloads to Lakehouse with minimal disruption.
  • Leverage Claude or Copilot to accelerate development, tests, and docs.
  • Write clean Python and SQL; ensure code quality via reviews.
  • Optimize Spark jobs and Delta tables for performance and cost.
  • Implement data quality, lineage, and governance controls.
  • Troubleshoot pipeline failures and production incidents.
  • Collaborate with DevOps, platform, and analytics teams on observability and security.

Conocimientos

Databricks
Apache Spark
PySpark
Delta Lake
SQL
Python
Structured Streaming
Kafka
Airflow
Azure Data Factory
dbt
Unity Catalog
Terraform
Cloud data architectures

Herramientas

Databricks
Airflow
Azure Data Factory
Terraform

Descripción del empleo

AgileEngine is seeking a Middle Data Engineer to modernize a 15-year data warehouse into a governed Databricks Lakehouse. You will build batch and streaming pipelines with PySpark and Delta Lake, using a medallion architecture for analytics, reporting and ML consumption.

Responsibilities include migrating legacy workloads, implementing data quality and governance, and collaborating with DevOps and analytics teams. Strong SQL/Python, Spark, and cloud architecture experience are required.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Mid Data Engineer - Databricks Lakehouse (Remote)
Mid Data Engineer - Databricks Lakehouse (Remote)

INGEPSY • Metropolitana

Presencial
COP 279.616.000 - 466.027.000
Growth opportunities
Remote work 100%
Learning budget
+2
Remote Data Engineer, Databricks Lakehouse
Remote Data Engineer, Databricks Lakehouse

AgileEngine • Perímetro Urbano Barranquilla

Presencial
COP 281.461.000 - 406.555.000
Growth opportunities
Competitive compensation
Remote work
+3
Data Engineer - Databricks Lakehouse & PySpark
Data Engineer - Databricks Lakehouse & PySpark

AgileEngine • Capital

Presencial
COP 189.167.000 - 315.278.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
+3
Data Engineer — Databricks Lakehouse (Remote)
Data Engineer — Databricks Lakehouse (Remote)

AgileEngine • Cartagena de Indias

Presencial
COP 205.918.000 - 300.957.000
Growth opportunities
Competitive compensation
100% remote
+5
Remote Data Engineer — Databricks Lakehouse & AI Pipelines
Remote Data Engineer — Databricks Lakehouse & AI Pipelines

INGEPSY • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Growth opportunities
Competitive compensation
Remote work
+3
Remote Data Engineer - Databricks Lakehouse & PySpark
Remote Data Engineer - Databricks Lakehouse & PySpark

AgileEngine • Centrosur

Presencial
COP 100.440.000 - 156.240.000
Growth without limits
Competitive compensation
Flexibility
+3
Senior Data Engineer – Databricks Lakehouse (Remote)
Senior Data Engineer – Databricks Lakehouse (Remote)

INGEPSY • Bogotá

Presencial
COP 217.479.000 - 341.753.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
Databricks Lakehouse Data Engineer — Remote
Databricks Lakehouse Data Engineer — Remote

AgileEngine, LLC. • Centrosur

Presencial
COP 6.000.000 - 9.000.000
Remote work
Flexible hours
Learning budget
+2
Data Engineer - Remote, Databricks Lakehouse & Pipelines
Data Engineer - Remote, Databricks Lakehouse & Pipelines

AgileEngine, LLC. • Metropolitana

Presencial
COP 90.000.000 - 140.000.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
Remote Data Engineer: Databricks Lakehouse & Pipelines
Remote Data Engineer: Databricks Lakehouse & Pipelines

AgileEngine • Metropolitana

Presencial
COP 283.751.000 - 409.862.000
Growth without limits
Competitive compensation
Flexibility
+3