Data Engineer

Sonatype

Norte

Presencial

COP 90.000.000 - 120.000.000

Jornada completa

Hace 9 días
Generador de candidaturas

Destaca para este puesto: genera un currículum y una carta de presentación adaptados en cuestión de un minuto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Parental leave
Diversity and inclusion working groups
Flexible working practices

Descripción de la vacante

Sonatype is seeking a data engineer to build and maintain reliable data pipelines and ETL/ELT workflows, partnering with teammates to deliver clean datasets for analytics and internal tooling.

You will develop data models, work with Spark, SQL, and distributed systems, and support cloud services like AWS. Strong English communication and a proactive, self-driven approach are essential for success in a fast-paced environment.

Formación

  • Bachelor's degree or equivalent practical experience in computer science, engineering, or a related field.
  • 2-4 years of experience in data engineering or backend data-related roles.
  • Strong programming skills in Java or Python and experience with SQL.
  • Experience with distributed data systems (Spark, Kafka, SQS).
  • Familiarity with NoSQL stores like Cassandra or HBase.
  • Proficient in English with strong communication; able to explain technical work to non-engineers.
  • Self-driven and capable of learning new technologies with minimal guidance.
  • Experience with debugging, code reviews, and CI/CD practices.
  • Open to using AI-assisted engineering tools in daily workflows.

Responsabilidades

  • Build and maintain reliable data pipelines and ETL/ELT workflows.
  • Develop and optimize data models for analytics and internal tools.
  • Collaborate with team to deliver clean, trusted datasets.
  • Support core data platform tools like Spark and AWS services (S3, SNS, SQS, ECS/Fargate, EMR).
  • Monitor data pipelines for quality, performance, and reliability.
  • Document workflows and contribute to test coverage and CI/CD processes.
  • Help shape data lakehouse architecture and platform roadmap.

Conocimientos

Java
Python
SQL
Spark
NoSQL
ETL/ELT
Communication

Educación

Bachelor's degree in CS/Engineering
Equivalent practical experience

Herramientas

Kafka
AWS
Databricks
dbt

Descripción del empleo

Sonatype is the software supply chain security company. We provide the world's best end-to-end software supply chain security solution, combining the only proactive protection against malicious open source, the only enterprise grade SBOM management and the leading open source dependency management platform. This empowers enterprises to create and maintain secure, quality, and innovative software at scale.

As founders of Nexus Repository and stewards of Maven Central, the world's largest repository of Java open-source software, we are software pioneers and our open source expertise is unmatched. We empower innovation with an unparalleled commitment to build faster, safer software and harness AI and data intelligence to mitigate risk, maximize efficiencies, and drive powerful software development.

More than 2,000 organizations, including 70% of the Fortune 100 and 15 million software developers, rely on Sonatype to optimize their software supply chains.

What you'll do:
  • Build and maintain reliable data pipelines and ETL/ELT workflows
  • Develop and optimize data models for analytics and internal tools
  • Work with team members to deliver clean, trusted datasets
  • Support core data platform tools like Spark and AWS (S3, SNS, SQS, ECS/Fargate, EMR)
  • Monitor data pipelines for quality, performance, and reliability
  • Write clear documentation and contribute to test coverage and CI/CD processes
  • Help shape our data lakehouse architecture and platform roadmap
What you'll bring:
  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience
  • 2-4 years of experience in data engineering or a backend data-related role
  • Strong skills in Java, Scala, or another backend programming language
  • Python (PySpark/pandas) skills
  • Experience with SQL and distributed data systems (e.g., Spark, Kafka, SQS)
  • Familiarity with NoSQL stores like Cassandra, HBase, or similar
  • Understanding of data modeling for analytics and reporting
  • Proficient in English with strong communication skills - able to explain or demo work to non-engineers
  • Self-driven - picks up new technologies and frameworks with little guidance
  • Debugs methodically; breaks complex problems into smaller steps
  • Open to feedback - iterates and improves through code reviewsUses AI-assisted engineering tools (Claude Code, Codex, Cursor, etc.) as part of daily workflow
  • Reliable internet connection to sustain video, audio, and screen sharing
It'd be great if you had:
  • Experience with dbt, Databricks, or real-time data pipelines
  • Familiarity with cloud infrastructure tools like Terraform or CloudFormation
  • Interest in data governance, ML pipelines, or compliance standards
  • Personal projects or open source contributions demonstrating initiative
Why you'll love working here:
  • Work on data that supports meaningful software security outcomes
  • Modern tools in a cloud-first, open-source-friendly environment
  • A team that values clarity, learning, and autonomy

At Sonatype, we value diversity and inclusivity. We offer perks such as parental leave, diversity and inclusion working groups, and flexible working practices to allow our employees to show up as their whole selves. We are an equal-opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. If you have a disability or special need that requires accommodation, please do not hesitate to let us know.

  • parental leave
  • diversity and inclusion working groups
  • flexible working practices
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Data Engineer
Data Engineer

Sonatype • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Parental leave
Diversity and inclusion groups
Flexible working practices
Java Product Support Engineer
Java Product Support Engineer

Sonatype • Bogotá

Presencial
COP 137.084.000 - 205.628.000
Parental leave
Diversity and inclusion working groups
Flexible working practices
+2
Data Engineer: Build Scalable Pipelines & Data Lakehouse
Data Engineer: Build Scalable Pipelines & Data Lakehouse

Sonatype Inc • Colombia

A distancia
COP 60.000.000 - 120.000.000
Parental leave
Diversity and Inclusion groups
Flexible working
Staff Software Engineer - Agentic First
Staff Software Engineer - Agentic First

Sonatype • Colombia

Presencial
COP 380.156.000 - 633.593.000
Parental leave
Diversity & inclusion groups
Flexible working
Data Engineer - Build Scalable Pipelines & Data Lakehouse
Data Engineer - Build Scalable Pipelines & Data Lakehouse

Sonatype • Norte

Presencial
COP 90.000.000 - 120.000.000
Parental leave
Diversity and inclusion working groups
Flexible working practices
Renewal Specialist - SaaS Contracts
Renewal Specialist - SaaS Contracts

Sonatype • Bogotá

Presencial
COP 60.000.000 - 90.000.000
Parental leave
Diversity & Inclusion programs
Flexible working practices
Data Engineer - Build Scalable Data Pipelines & Lakehouse
Data Engineer - Build Scalable Data Pipelines & Lakehouse

Sonatype • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Parental leave
Diversity and inclusion groups
Flexible working practices
Senior Data Engineer ID71670
Senior Data Engineer ID71670

AgileEngine • Bogotá

Híbrido
COP 281.171.000 - 406.136.000
Professional growth
Competitive compensation
Exciting projects
+1
Data Engineer Sr- Remote Latam – ID #00199
Data Engineer Sr- Remote Latam – ID #00199

Werben HR • Bogotá

A distancia
COP 100.000.000 - 120.000.000
Senior Staff Engineer - Data Engineer
Senior Staff Engineer - Data Engineer

Nagarro • Colombia

A distancia
COP 172.526.000 - 249.205.000