Data Engineer LATAM

Onebeat

Rio de Janeiro

Presencial

BRL 120 000 - 240 000

Tempo integral

há 46 horas
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Uma candidatura feita para esta oferta — um currículo e uma carta de apresentação personalizados que vão ao encontro do anúncio.

Ultrapassa os filtros ATS

Resumo da oferta

Onebeat is seeking an experienced Data Engineer to design, develop, and maintain an open-source data platform. You will work on a Data-Lakehouse using S3, Apache Iceberg, and ClickHouse, building scalable ETL/ELT pipelines and orchestration with Temporal.

You will collaborate with data scientists, analysts, and engineers to deliver reliable data solutions, enforce governance, and drive data quality across the platform in a batch-processing environment.

Qualificações

  • 3+ years of experience as a Data Engineer or similar data infrastructure role.
  • Strong proficiency in SQL and hands-on experience with data modeling.
  • Experience with data lake/lakehouse architectures (e.g., Apache Iceberg, S3, or similar).
  • Experience with analytical / columnar databases (e.g., ClickHouse or similar).
  • Experience building and orchestrating ETL/ELT pipelines (e.g., Temporal, Airflow, or similar).
  • Strong programming skills in Python and/or Scala/Java.
  • Experience working within a microservices architecture and cloud environments (AWS preferred).
  • Self-motivated, strong multitasking skills, and a demonstrated team player.
  • Excellent communication skills and the ability to work both independently and collaboratively.
  • Hands-on experience with Apache Spark (or similar technologies) for large-scale data processing.
  • Professional proficiency in written and spoken English.
  • Note: this role is focused on batch data processing (not real-time streaming).

Responsabilidades

  • Develop a scalable data platform integrating multiple sources for easy access.
  • Design and enhance data tools (orchestration, governance, Data-Lakehouse, BI, etc.).
  • Ensure smooth operation of data systems for analysts, scientists, and engineers.
  • Optimize data pipelines (ingestion, processing, and output) in a microservices environment.
  • Build, maintain, and monitor ETL/ELT processes and orchestrate workflows using Temporal.
  • Troubleshoot and improve the performance, scalability, and reliability of the data infrastructure (S3, Apache Iceberg, ClickHouse).
  • Collaborate cross-functionally with data scientists, analysts, and backend engineers to understand data needs and deliver solutions.
  • Implement and champion data quality, governance, and security best practices across the platform.

Conhecimentos

SQL proficiency
Data modeling
Python
Scala/Java
ETL/ELT pipelines
Temporal/Airflow
Cloud AWS
Team collaboration
Communication
Spark

Ferramentas

Apache Iceberg
ClickHouse
S3
Temporal
Airflow
Spark

Descrição da oferta de emprego

Description

We are seeking an experienced Data Engineer. The ideal candidate is self-motivated, a multitasker, and a demonstrated team player. You will be responsible for designing, developing, managing, and maintaining our open-source data platform, including our Data-Lakehouse (S3, Apache Iceberg, and ClickHouse), ETL processes, and orchestration tool (Temporal Workflow).

What You Will Do
  • Develop a scalable data platform integrating multiple sources for easy access.
  • Design and enhance data tools (orchestration, governance, Data-Lakehouse, BI, etc.).
  • Ensure smooth operation of data systems for analysts, scientists, and engineers.
  • Optimize data pipelines (ingestion, processing, and output) in a microservices environment.
  • Build, maintain, and monitor ETL/ELT processes and orchestrate workflows using Temporal.
  • Troubleshoot and improve the performance, scalability, and reliability of the data infrastructure (S3, Apache Iceberg, ClickHouse).
  • Collaborate cross-functionally with data scientists, analysts, and backend engineers to understand data needs and deliver solutions.
  • Implement and champion data quality, governance, and security best practices across the platform.
Requirements
  • 3+ years of experience as a Data Engineer or in a similar data infrastructure role.
  • Strong proficiency in SQL and hands-on experience with data modeling.
  • Experience with data lake/lakehouse architectures (e.g., Apache Iceberg, S3, or similar).
  • Experience with analytical / columnar databases (e.g., ClickHouse or similar).
  • Experience building and orchestrating ETL/ELT pipelines (e.g., Temporal, Airflow, or similar).
  • Strong programming skills in Python and/or Scala/Java.
  • Experience working within a microservices architecture and cloud environments (AWS preferred).
  • Self-motivated, strong multitasking skills, and a demonstrated team player.
  • Excellent communication skills and the ability to work both independently and collaboratively.
  • Hands-on experience with Apache Spark (or similar technologies) for large-scale data processing.
  • Professional proficiency in written and spoken English.
  • Note: this role is focused on batch data processing (not real-time streaming).
Nice to Have
  • Experience working with and contributing to open-source data platforms and tools.
  • Familiarity with BI and visualization tools (e.g., Superset, Looker, Tableau, Metabase, or similar).
  • Experience with containerization and orchestration (Docker, Kubernetes).
  • Experience with infrastructure-as-code and CI/CD practices.
  • Experience with AWS EMR and running Apache Spark workloads in a cloud environment.
  • Experience leveraging AI-assisted development tools (e.g., GitHub Copilot, Cursor, or similar) to boost engineering productivity.
Description

We are seeking an experienced Data Engineer. The ideal candidate is self-motivated, a multitasker, and a demonstrated team player. You will be responsible for designing, developing, managing, and maintaining our open-source data platform, including our Data-Lakehouse (S3, Apache Iceberg, and ClickHouse), ETL processes, and orchestration tool (Temporal Workflow).

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Data Engineer
Data Engineer

Microtalent is becoming INSPYR Global Solutions • Brasil

Presencial
Data Engineer (Tech-lead) - 1894
Data Engineer (Tech-lead) - 1894

In All Media Inc • Brasil

Teletrabalho
BRL 361 000 - 465 000
Lead Data DevOps Engineer
Lead Data DevOps Engineer

EPAM Systems • Brasil

Presencial
BRL 220 000 - 320 000
Lead Data Engineer
Lead Data Engineer

EPAM Systems • Brasil

Presencial
BRL 180 000 - 300 000
Databricks certification
Data Engineer - Senior - BR
Data Engineer - Senior - BR

Imagemaker • São Paulo

Presencial
BRL 120 000 - 150 000
Data Engineer - Remote, Latin America
Data Engineer - Remote, Latin America

Bluelight • Belo Horizonte

Teletrabalho
Competitive salary and bonuses
Generous paid-time-off policy
Technology / Office stipend
+4
Lead Data Software Engineer (Java+AWS)
Lead Data Software Engineer (Java+AWS)

EPAM Systems • Brasil

Presencial
BRL 180 000 - 240 000
Senior Data Engineer
Senior Data Engineer

Athenaworks • Brasil

Presencial
BRL 618 000 - 824 000
Payment in USD
Flexible work schedule
Learning Budget
+2
Data Engineer (Lead) ID52236
Data Engineer (Lead) ID52236

AgileEngine • Belo Horizonte

Híbrido
BRL 421 000 - 580 000
Mentorship and personalized growth roadmaps
Competitive compensation with budgets for fitness and education
Exciting projects with Fortune 500 brands
+1
Data Platform Engineer
Data Platform Engineer

Avenue Code • Brasil

Presencial
BRL 120 000 - 240 000