Databricks Data Engineer: Lakehouse AI Pipelines Architect

AgileEngine, LLC.

Sur

Híbrido

COP 280.260.000 - 435.961.000

Jornada completa

Hace 4 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

No envíes un currículum genérico: crea un currículum y una carta de presentación adaptados a este puesto concreto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Professional growth
Competitive USD-based compensation
A selection of exciting projects
Flextime

Descripción de la vacante

AgileEngine is seeking a Data Engineer to build batch and streaming pipelines on Databricks using PySpark and Delta Lake. This role migrates legacy data workloads to a governed Lakehouse, modeling a medallion architecture that powers analytics and AI use cases.

The role emphasizes SQL and Python, data governance with Unity Catalog, and collaboration with DevOps and analytics engineers. Flexible work options are available.

Formación

  • 4+ years of professional data engineering experience with Spark and cloud data architectures.
  • Strong hands-on experience with Databricks, Spark (PySpark), and Delta Lake.
  • Advanced SQL and Python with data modeling skills for Lakehouse patterns.
  • Experience with streaming ingestion (Structured Streaming, Auto Loader, Kafka, Event Hubs).
  • Orchestrating data workflows (Databricks Workflows, Airflow, Azure Data Factory).
  • Experience migrating legacy ETL and data warehouses to Lakehouse, data hygiene.
  • Strong communication, mentoring, and collaboration skills; familiar with Unity Catalog and PII handling.
  • Experience with dbt or equivalent transformation framework; secure coding practices.

Responsabilidades

  • Design, build, and operate batch and streaming data pipelines on Databricks.
  • Model and maintain a medallion architecture for analytics, reporting, and ML.
  • Migrate legacy ETL and data warehouse workloads onto the Lakehouse with parity.
  • Use Claude or GitHub Copilot to accelerate development, tests, docs, and prototyping.
  • Write clean Python and SQL; maintain quality via code reviews.
  • Optimize Spark jobs and Delta tables for performance and cost.
  • Implement data quality, lineage, governance using Unity Catalog.
  • Debug, troubleshoot, and resolve production data issues.
  • Participate in Agile delivery practices; collaborate with DevOps and analytics engineers.
  • Collaborate with security/compliance teams on observability and governance.

Conocimientos

Data engineering
Apache Spark
PySpark
Delta Lake
SQL
Python
Structured Streaming
Kafka
Event Hubs
Data modeling
Lakehouse
Mentoring juniors
Data governance
PII handling
dbt
Security best practices
Problem solving
Collaboration

Herramientas

Databricks
Databricks Workflows
Airflow
Azure Data Factory
Terraform
PostgreSQL

Descripción del empleo

AgileEngine is seeking a Data Engineer to build batch and streaming pipelines on Databricks using PySpark and Delta Lake. This role migrates legacy data workloads to a governed Lakehouse, modeling a medallion architecture that powers analytics and AI use cases.

The role emphasizes SQL and Python, data governance with Unity Catalog, and collaboration with DevOps and analytics engineers. Flexible work options are available.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Databricks Data Engineer: Lakehouse & AI Analytics
Databricks Data Engineer: Lakehouse & AI Analytics

INGEPSY • Perímetro Urbano Barranquilla

Híbrido
COP 280.260.000 - 404.820.000
Professional growth
Competitive USD-based compensation
A selection of exciting projects
+1
Databricks Data Engineer - Remote-Friendly Lakehouse & Pipelines
Databricks Data Engineer - Remote-Friendly Lakehouse & Pipelines

AgileEngine, LLC. • Perímetro Urbano Barranquilla

Híbrido
COP 155.700.000 - 342.540.000
Professional growth
Competitive compensation (USD-based)
A selection of exciting projects
+1
Databricks Data Engineer — Flexible, High-Impact Analytics
Databricks Data Engineer — Flexible, High-Impact Analytics

AgileEngine, LLC. • Pereira

Híbrido
COP 280.260.000 - 404.820.000
Professional growth
Competitive USD-based compensation
Exciting projects with top-tier brands
+1
Senior Databricks Data Engineer - Lakehouse & ML Pipelines
Senior Databricks Data Engineer - Lakehouse & ML Pipelines

INGEPSY • Metropolitana

Presencial
COP 280.260.000 - 435.961.000
Professional growth
Competitive compensation
Curated projects
+1
Senior Databricks Engineer - Lakehouse & Data Platforms
Senior Databricks Engineer - Lakehouse & Data Platforms

INGEPSY • Sur

Presencial
COP 280.260.000 - 467.101.000
Competitive USD-based compensation
Flextime
Professional growth
+1
Senior Databricks Engineer: Flexible Data Pipelines
Senior Databricks Engineer: Flexible Data Pipelines

AgileEngine, LLC. • Pereira

Híbrido
COP 100.000.000 - 150.000.000
Flextime
Competitive USD-based compensation
Professional growth & mentorship
+1
Senior Databricks Engineer: Lakehouse Data Platform Lead
Senior Databricks Engineer: Lakehouse Data Platform Lead

INGEPSY • Sucre

Presencial
COP 120.000.000 - 160.000.000
Professional growth
Competitive USD-based compensation
A selection of exciting projects
+1
Senior Databricks & Lakehouse Engineer
Senior Databricks & Lakehouse Engineer

INGEPSY • Capital

Híbrido
COP 280.260.000 - 404.820.000
Professional growth
Competitive USD compensation
A selection of exciting projects
+1
Databricks Engineer: Lakehouse & Streaming Architect
Databricks Engineer: Lakehouse & Streaming Architect

INGEPSY • Risaralda

Presencial
COP 280.260.000 - 435.961.000
Professional growth
Competitive USD-based compensation
A selection of exciting projects
+1
Senior Databricks Engineer | Hybrid Data Platform Architect
Senior Databricks Engineer | Hybrid Data Platform Architect

INGEPSY • Bogotá ciudad

Híbrido
COP 110.000.000 - 150.000.000
Professional growth
Competitive compensation
A selection of exciting projects
+1