Senior Data Platform Engineer

Solvd, Inc.

Antioquia

Presencial

COP 120.000.000 - 180.000.000

Jornada completa

Hace 3 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Una candidatura hecha para este puesto de trabajo: un currículum y una carta de presentación adaptados que responden directamente a la oferta.

Supera los filtros ATS

Descripción de la vacante

Solvd Inc. is seeking a Senior Data Platform Engineer to join an engagement building data pipelines and lakehouse architecture for a scientific data cloud platform — purpose-built for biopharma.

You will turn raw instrument output from lab equipment into harmonized scientific data that accelerates research outcomes, working end-to-end with customers. This role focuses on prototyping pipelines, building data models, and shipping production-grade solutions.

Formación

  • 5+ years of professional experience in Python and SQL.
  • Hands-on experience with Databricks Lakehouse on AWS.
  • Strong working knowledge of AWS Redshift — hands-on experience required.
  • Experience across the AWS data stack: ECS, S3, Athena, RDS.
  • Proficiency with ETL/ELT pipeline development using Airflow and Python.
  • Experience with relational databases — MySQL, MariaDB, Aurora, PostgreSQL, MS SQL Server.
  • Familiarity with key-value and non-relational databases.
  • Excellent communication skills and the confidence to take ownership of project delivery end-to-end.
  • Able to manage multiple simultaneous projects without losing quality or attention to detail.
  • Genuine curiosity — learning new scientific domains and instrument ecosystems continuously.

Responsabilidades

  • Own, prototype, implement, and deploy DataBricks Lakehouse pipelines on AWS.
  • Research and prototype data acquisition strategies for scientific lab instrumentation.
  • Build file parsers for instrument output files across formats — .xlsx, .pdf, .txt, .raw, .fid, and vendor-specific binaries.
  • Design and build data models, Python data pipelines, unit tests, integration tests, and utility functions.
  • Work directly with customers to validate that solutions meet their requirements and solve real scientific needs.
  • Facilitate internal project post-mortems to identify and apply improvements across engagements.

Conocimientos

Python
SQL
Databricks
AWS
Airflow
ETL/ELT
Relational DBs
NoSQL
Communication

Educación

Bachelor's or master's in CS/related field

Herramientas

S3
Athena
ECS
RDS

Descripción del empleo

Solvd Inc. is a rapidly growing AI-native consulting and technology services firm delivering enterprise transformation across cloud, data, software engineering, and artificial intelligence. We work with industry-leading organizations to design, build, and operationalize technology solutions that drive measurable business outcomes.

Following the acquisition of Tooploox, a premier AI and product development company, Solvd now offers true end-to-end delivery—from strategic advisory and solution design to custom AI development and enterprise-scale implementation. Our capability centers combine deep technical expertise, proven delivery methodologies, and sector-specific knowledge to address complex business challenges quickly and effectively.

We are looking for a Senior Data Platform Engineer to join an engagement building data pipelines and lakehouse architecture for a scientific data cloud platform — purpose-built for biopharma. You'll be working at the intersection of data engineering and life sciences, turning raw instrument output from lab equipment into harmonized, actionable scientific data that accelerates research outcomes.

This is hands-on, end-to-end engineering work: prototyping pipelines, parsing proprietary instrument file formats, designing data models, and shipping production-grade solutions directly with the customer.

What you'll do

Own, prototype, implement, and deploy DataBricks Lakehouse pipelines on AWS.

Research and prototype data acquisition strategies for scientific lab instrumentation.

Build file parsers for instrument output files across a wide range of formats — .xlsx, .pdf, .txt, .raw, .fid, and vendor-specific binaries.

Design and build data models, Python data pipelines, unit tests, integration tests, and utility functions.

Work directly with customers to validate that solutions meet their requirements and solve real scientific needs.

Facilitate internal project post-mortems to identify and apply improvements across engagements.

What you bring

5+ years of professional experience in Python and SQL.

Hands-on experience with Databricks Lakehouse architecture on AWS.

Strong working knowledge of AWS Redshift — hands-on experience required.

Experience across the AWS data stack: ECS, S3, Athena, RDS.

Proficiency with ETL/ELT pipeline development using Airflow and Python.

Experience with relational databases — MySQL, MariaDB, Aurora, PostgreSQL, MS SQL Server.

Familiarity with key-value and non-relational databases.

Excellent communication skills and the confidence to take ownership of project delivery end-to-end.

Able to manage multiple simultaneous projects without losing quality or attention to detail.

Genuine curiosity — this role requires learning new scientific domains and instrument ecosystems continuously.

Nice to have

Elasticsearch experience.

Background in or exposure to scientific instrumentation, lab workflows, or biopharma data.

Graduate degree in Chemistry, Biology, Computer Science, Statistics, Public Health, or a related field.

Experience parsing proprietary binary file formats from lab instruments.

Shape real-world AI-driven projects across key industries, working with clients from startup innovation to enterprise transformation.

Be part of a global team with equal opportunities for collaboration across continents and cultures.

Thrive in an inclusive environment that prioritizes continuous learning, innovation, and ethical AI standards.

Ready to make an impact?

An AI-first advisory and digital engineering firm that guides brands through digital transformation and delivers measurable impact | Solvd

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Databricks MLOps Engineer
Databricks MLOps Engineer

Solvd, Inc. • Colombia

Presencial
COP 281.171.000 - 406.136.000
Senior Data Platform Engineer — Life Sciences Lakehouse
Senior Data Platform Engineer — Life Sciences Lakehouse

Solvd, Inc. • Antioquia

Presencial
COP 120.000.000 - 180.000.000
Senior Data Lead Engineer
Senior Data Lead Engineer

UST España & Latam • Colombia

Presencial
COP 256.166.000 - 329.357.000
Senior Data Engineer
Senior Data Engineer

Publicis Groupe Holdings B.V • Bogotá

Presencial
COP 219.042.000 - 292.057.000
Senior Data Engineer
Senior Data Engineer

MPS Group LLC • Bogotá

Presencial
COP 120.000.000 - 180.000.000
1104 | Senior Data (Databricks) Engineer
1104 | Senior Data (Databricks) Engineer

Intetics • Colombia

Presencial
COP 108.000.000 - 180.000.000
Full Stack Software Engineer Id85839
Full Stack Software Engineer Id85839

INGEPSY • Risaralda

Presencial
COP 279.616.000 - 466.027.000
Staff Data Engineer
Staff Data Engineer

Robots & Pencils • Colombia

Híbrido
COP 150.000.000 - 250.000.000
Senior Data Engineer
Senior Data Engineer

Bespoke Labs • Colombia

A distancia
COP 387.341.000 - 602.531.000
Data Engineer
Data Engineer

Blanc Labs • Colombia

Presencial
COP 221.905.000 - 348.708.000