Data Engineer: Build Scalable Pipelines & Data Lakehouse

Sonatype Inc

Colombia

A distancia

COP 60.000.000 - 120.000.000

Jornada completa

Hace 10 días
Generador de candidaturas

Transforma esta oferta en una entrevista: un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Parental leave
Diversity and Inclusion groups
Flexible working

Descripción de la vacante

Sonatype is seeking a data engineer to build and maintain reliable data pipelines, develop analytics-ready data models, and ensure data quality across platforms. You will work with Spark, AWS services, and NoSQL stores while applying SQL and Python/Java/Scala skills across collaborative teams.

The role emphasizes CI/CD, documentation, and opportunities to influence our data lakehouse architecture and platform roadmap, in a cloud-first environment that values autonomy and learning.

Formación

  • Bachelor's degree in Computer Science, Engineering, or related field.
  • 2–4 years of experience in data engineering or backend data roles.
  • Strong skills in Java, Scala, or another backend programming language.
  • Python (PySpark/pandas) skills.
  • Experience with SQL and distributed data systems (e.g., Spark, Kafka, SQS).
  • Familiarity with NoSQL stores like Cassandra, HBase, or similar.
  • Understanding of data modeling for analytics and reporting.
  • Proficient in English with strong communication skills—able to explain or demo work to non-engineers.
  • Self-driven — picks up new technologies and frameworks with little guidance.
  • Debugs methodically; breaks complex problems into smaller steps.
  • Open to feedback — iterates and improves through code reviews.
  • Uses AI-assisted engineering tools (Claude Code, Codex, Cursor, etc.) as part of daily workflow.
  • Reliable internet connection to sustain video, audio, and screen sharing.

Responsabilidades

  • Build and maintain reliable data pipelines and ETL/ELT workflows.
  • Develop and optimize data models for analytics and internal tools.
  • Work with team members to deliver clean, trusted datasets.
  • Support core data platform tools like Spark and AWS (S3, SNS, SQS, ECS/Fargate, EMR).
  • Monitor data pipelines for quality, performance, and reliability.
  • Write clear documentation and contribute to test coverage and CI/CD processes.
  • Help shape our data lakehouse architecture and platform roadmap.

Conocimientos

Java
Scala
Python
SQL
Spark
Kafka
SQS
NoSQL
English
Communication
Problem-solving

Educación

Bachelor's degree in Computer Science/Engineering or related field

Herramientas

dbt
Databricks
Terraform
CloudFormation

Descripción del empleo

Sonatype is seeking a data engineer to build and maintain reliable data pipelines, develop analytics-ready data models, and ensure data quality across platforms. You will work with Spark, AWS services, and NoSQL stores while applying SQL and Python/Java/Scala skills across collaborative teams.

The role emphasizes CI/CD, documentation, and opportunities to influence our data lakehouse architecture and platform roadmap, in a cloud-first environment that values autonomy and learning.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Data Engineer - Build Scalable Data Pipelines & Lakehouse
Data Engineer - Build Scalable Data Pipelines & Lakehouse

Sonatype • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Parental leave
Diversity and inclusion groups
Flexible working practices
Data Engineer - Build Scalable Pipelines & Data Lakehouse
Data Engineer - Build Scalable Pipelines & Data Lakehouse

Sonatype • Norte

Presencial
COP 90.000.000 - 120.000.000
Parental leave
Diversity and inclusion working groups
Flexible working practices
Data Engineer
Data Engineer

Sonatype • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Parental leave
Diversity and inclusion groups
Flexible working practices
Data Engineer
Data Engineer

Sonatype • Norte

Presencial
COP 90.000.000 - 120.000.000
Parental leave
Diversity and inclusion working groups
Flexible working practices
Senior Data Engineer: Lead Scalable Pipelines & Data Platform
Senior Data Engineer: Lead Scalable Pipelines & Data Platform

Lean Tech, a Lean Solutions Group Division • Medellín

Presencial
COP 100.000.000 - 160.000.000
Data Engineer - Databricks Lakehouse & AI-Driven Pipelines
Data Engineer - Databricks Lakehouse & AI-Driven Pipelines

AgileEngine • Pereira

Presencial
COP 84.000.000 - 126.000.000
Growth budget
Competitive pay
Remote work options
+3
Senior Data Engineer: Build Scalable Pipelines & Analytics
Senior Data Engineer: Build Scalable Pipelines & Analytics

DevSavant Inc. • Colombia

A distancia
COP 150.000.000 - 230.000.000
Data Engineer - Databricks Lakehouse & Pipelines
Data Engineer - Databricks Lakehouse & Pipelines

AgileEngine, LLC. • Perímetro Urbano Barranquilla

Presencial
COP 289.119.000 - 417.617.000
Growth without limits
Competitive compensation
Remote work with flexible hours
+3
Data Engineer: Build Scalable Pipelines & Trusted Data
Data Engineer: Build Scalable Pipelines & Trusted Data

StartX • Bogotá

Presencial
COP 90.000.000 - 130.000.000
Health benefits
Life Insurance
Indefinite-term contract
+5
Senior Databricks Data Engineer: Lakehouse & AI Pipelines
Senior Databricks Data Engineer: Lakehouse & AI Pipelines

AgileEngine, LLC. • Cartagena de Indias

Presencial
COP 385.493.000 - 578.239.000
Growth without limits
Competitive compensation
Flexibility: remote work with flexible
+3