Data Engineer — Databricks Lakehouse (Remote)

AgileEngine

Cartagena de Indias

Presencial

COP 205.918.000 - 300.957.000

Jornada completa

Hace 9 días
Generador de candidaturas

No envíes un currículum genérico: crea un currículum y una carta de presentación adaptados a este puesto concreto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Growth opportunities
Competitive compensation
100% remote
Annual learning budget
Flexible hours
Modern projects
Collaborative culture
Well-being programs

Descripción de la vacante

AgileEngine is seeking a Middle Data Engineer to modernize a 15-year-old data warehouse into a governed Databricks Lakehouse. You will design batch and streaming pipelines with PySpark and Delta Lake, and implement a medallion architecture across bronze, silver, and gold layers.

This role also uses AI tools like Claude and GitHub Copilot to speed up development, emphasizes data quality, governance with Unity Catalog, and collaboration with DevOps, platform, and analytics engineers.

Formación

  • 3+ years of professional data engineering experience
  • Proficient with Spark and cloud-based data architectures
  • Strong SQL and Python skills
  • Experience with streaming using Structured Streaming, Auto Loader, Kafka or Event Hubs
  • Experience with Databricks Workflows, Airflow or Azure Data Factory
  • Experience with legacy platform migrations and ETL modernization
  • Strong problem-solving and communication skills
  • Familiarity with Unity Catalog and data governance
  • Experience with dbt or equivalent transformation framework
  • Upper-intermediate English

Responsabilidades

  • Design, build, and operate batch and streaming data pipelines on Databricks using PySpark, Delta Lake, and Databricks Workflows.
  • Model and maintain a medallion (bronze/silver/gold) architecture for analytics, reporting, and ML consumers.
  • Migrate legacy ETL and data warehouse workloads onto the Lakehouse with validated data parity and minimal disruption.
  • Use Claude or GitHub Copilot to accelerate development, write tests, and create documentation.
  • Write clean Python and SQL; ensure quality via code reviews and tests.
  • Optimize Spark jobs and Delta tables for performance and cost.
  • Implement data quality, lineage, and governance with Unity Catalog.
  • Debug and resolve pipeline failures and production incidents.
  • Collaborate with DevOps, platform, and analytics engineers on observability, security, and compliance.

Conocimientos

Data engineering
Apache Spark
Databricks
Python
SQL

Herramientas

Databricks
Apache Spark (PySpark)
Delta Lake

Descripción del empleo

AgileEngine is seeking a Middle Data Engineer to modernize a 15-year-old data warehouse into a governed Databricks Lakehouse. You will design batch and streaming pipelines with PySpark and Delta Lake, and implement a medallion architecture across bronze, silver, and gold layers.

This role also uses AI tools like Claude and GitHub Copilot to speed up development, emphasizes data quality, governance with Unity Catalog, and collaboration with DevOps, platform, and analytics engineers.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Databricks Lakehouse Data Engineer — Remote
Databricks Lakehouse Data Engineer — Remote

AgileEngine, LLC. • Centrosur

Presencial
COP 6.000.000 - 9.000.000
Remote work
Flexible hours
Learning budget
+2
Data Engineer - Remote, Databricks Lakehouse & Pipelines
Data Engineer - Remote, Databricks Lakehouse & Pipelines

AgileEngine, LLC. • Metropolitana

Presencial
COP 90.000.000 - 140.000.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
Remote Data Engineer - Databricks Lakehouse & PySpark
Remote Data Engineer - Databricks Lakehouse & PySpark

AgileEngine • Centrosur

Presencial
COP 100.440.000 - 156.240.000
Growth without limits
Competitive compensation
Flexibility
+3
Remote Data Engineer - Databricks Lakehouse & AI Pipelines
Remote Data Engineer - Databricks Lakehouse & AI Pipelines

AgileEngine, LLC. • Bogotá

Presencial
COP 90.000.000 - 150.000.000
Growth budget
Competitive compensation
Remote work
+3
Data Engineer - Databricks Lakehouse & PySpark
Data Engineer - Databricks Lakehouse & PySpark

AgileEngine • Capital

Presencial
COP 189.167.000 - 315.278.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
+3
Databricks Data Engineer — Lakehouse, PySpark & AI Tools
Databricks Data Engineer — Lakehouse, PySpark & AI Tools

AgileEngine, LLC. • Medellín

Presencial
COP 289.119.000 - 417.617.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
+3
Remote Data Engineer — Databricks Lakehouse & Pipelines
Remote Data Engineer — Databricks Lakehouse & Pipelines

AgileEngine • Bogotá

Presencial
COP 189.167.000 - 283.751.000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
Remote Data Engineer: Databricks Lakehouse & Pipelines
Remote Data Engineer: Databricks Lakehouse & Pipelines

AgileEngine • Metropolitana

Presencial
COP 283.751.000 - 409.862.000
Growth without limits
Competitive compensation
Flexibility
+3
Data Engineer - Databricks Lakehouse & Pipelines
Data Engineer - Databricks Lakehouse & Pipelines

AgileEngine, LLC. • Perímetro Urbano Barranquilla

Presencial
COP 289.119.000 - 417.617.000
Growth without limits
Competitive compensation
Remote work with flexible hours
+3
Data Engineer - Databricks Lakehouse & AI-Driven Pipelines
Data Engineer - Databricks Lakehouse & AI-Driven Pipelines

AgileEngine • Pereira

Presencial
COP 84.000.000 - 126.000.000
Growth budget
Competitive pay
Remote work options
+3