Data Engineer (DBT + Spark + Argo) (Remote - Latam)

Jobgether

Brasil

Teletrabalho

BRL 435 492 - 653 239

Tempo integral

14 dias+
Gerador de candidaturas

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Contractor agreement with payment in USD
100% remote work within LATAM
Access to English classes and professional learning platforms
Referral program and growth opportunities

Resumo da oferta

A technology partnership platform is seeking a skilled Data Engineer to work remotely in Latin America. This role involves transforming legacy SQL pipelines into scalable DBT architectures, utilizing Spark for processing, and Argo for workflow orchestration, with a focus on healthcare data solutions. Candidates should have experience with DBT and Spark SQL, as well as strong Python programming skills. The position offers competitive payment in USD and growth opportunities in state-of-the-art projects.

Qualificações

  • Strong experience with DBT for data modeling, testing, and deployment.
  • Hands-on proficiency in Spark SQL, including performance tuning.
  • Solid programming skills in Python for automation and data manipulation.

Responsabilidades

  • Transform legacy SQL pipelines into modular DBT architectures.
  • Build high-performance data transformation pipelines.
  • Design automated workflows using Argo Workflows.

Conhecimentos

DBT for data modeling
Spark SQL
Python skills for automation
Jinja templating
Apache Hudi, Parquet, Iceberg
Argo Workflows
AWS S3 storage
ElasticSearch
Healthcare data standards
Agile work environments

Ferramentas

AWS Glue
Docker
Kubernetes

Descrição da oferta de emprego

Get AI-powered advice on this job and more exclusive features.

This position is posted by Jobgether on behalf of a partner company. We are currently looking for a Data Engineer (DBT + Spark + Argo) in Latin America.

We are seeking a highly skilled Data Engineer to join a remote-first, collaborative team driving the modernization of large-scale data platforms in the healthcare sector. In this role, you will work on transforming legacy SQL pipelines into modular, scalable, and testable DBT architectures, leveraging Spark for high-performance processing and Argo for workflow orchestration. You will implement modern lakehouse solutions, optimize storage and querying strategies, and enable real-time analytics with ElasticSearch. This position offers the chance to contribute to a cutting-edge, cloud-native data environment, working closely with cross-functional teams to deliver reliable, impactful data solutions.

Accountabilities
  • Translate legacy T-SQL logic into modular, scalable DBT models powered by Spark SQL
  • Build reusable, high-performance data transformation pipelines
  • Develop testing frameworks to ensure data accuracy and integrity within DBT workflows
  • Design and orchestrate automated workflows using Argo Workflows and CI/CD pipelines with Argo CD
  • Manage reference datasets and mock data (e.g., ICD-10, CPT), maintaining version control and governance
  • Implement efficient storage and query strategies using Apache Hudi, Parquet, and Iceberg
  • Integrate ElasticSearch for analytics through APIs and pipelines supporting indexing and querying
  • Collaborate with DevOps teams to optimize cloud storage, enforce security, and ensure compliance
  • Participate in Agile squads, contributing to planning, estimation, and sprint reviews
Requirements
  • Strong experience with DBT for data modeling, testing, and deployment
  • Hands-on proficiency in Spark SQL, including performance tuning
  • Solid programming skills in Python for automation and data manipulation
  • Familiarity with Jinja templating to build reusable DBT components
  • Practical experience with data lake formats: Apache Hudi, Parquet, Iceberg
  • Expertise in Argo Workflows and CI/CD integration with Argo CD
  • Deep understanding of AWS S3 storage, performance tuning, and cost optimization
  • Experience with ElasticSearch for indexing and querying structured/unstructured data
  • Knowledge of healthcare data standards (e.g., ICD-10, CPT)
  • Ability to work cross-functionally in Agile environments
  • Nice to have: Experience with Docker, Kubernetes, cloud-native data tools (AWS Glue, Databricks, EMR), CI/CD automation, data compliance standards (HIPAA, SOC2), or contributions to open-source DBT/Spark projects
Benefits
  • Contractor agreement with payment in USD
  • 100% remote work within LATAM
  • Observance of local public holidays
  • Access to English classes and professional learning platforms
  • Referral program and other growth opportunities
  • Exposure to cutting-edge data engineering projects in a cloud-native environment

Jobgether is a Talent Matching Platform that partners with companies worldwide to efficiently connect top talent with the right opportunities through AI-driven job matching.

When you apply, your profile goes through our AI-powered screening process designed to identify top talent efficiently and fairly.

  • 🔍 Our AI thoroughly analyzes your CV and LinkedIn profile, evaluating your skills, experience, and achievements
  • 📊 It compares your profile against the job's core requirements and past success factors to calculate a match score
  • 🎯 The top 3 candidates with the highest match are automatically shortlisted
  • 🧠 When necessary, our human team may perform additional review to ensure no strong candidate is overlooked

The process is transparent, skills-based, and unbiased, focusing solely on your fit for the role. Once the shortlist is completed, it is shared with the hiring company, who then determines next steps such as interviews or additional assessments.

Thank you for your interest!

Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

Data Engineer - Remote Work
Data Engineer - Remote Work

BairesDev • São Paulo

Presencial
BRL 429 876 - 644 815
100% remote work
Excellent compensation in USD
Flexible hours
+3
Data Engineer - Remote Work
Data Engineer - Remote Work

BairesDev • São Paulo

Presencial
BRL 457 974 - 646 551
100% remote work
Excellent compensation in USD
Flexible hours
+3
Databricks Data Engineer (Apache Spark) - Remote work | REF#303795
Databricks Data Engineer (Apache Spark) - Remote work | REF#303795

BairesDev • Minas Gerais

Presencial
BRL 461 000 - 768 000
Remote work
USD or local currency pay
Home office setup
+4
Data Engineer (Tech-lead) - 1894
Data Engineer (Tech-lead) - 1894

In All Media Inc • Brasil

Presencial
BRL 361 570 - 464 876
Senior Data Engineer / Data Engineering Technical Lead
Senior Data Engineer / Data Engineering Technical Lead

Lever, Inc. • Brasil

Teletrabalho
BRL 180 000 - 300 000
Remote work opportunity
Modern data stack exposure
Technical leadership responsibilities
Senior Data Engineer 2 - Remote Work | REF#297238
Senior Data Engineer 2 - Remote Work | REF#297238

BairesDev • São Paulo

Presencial
BRL 614 000 - 922 000
100% remote
USD or local currency compensation
Hardware provided
+4
Data & Analytics Engineer (Remote - Latam)
Data & Analytics Engineer (Remote - Latam)

Jobgether • Brasil

Presencial
BRL 374 331 - 481 283
Fully remote work opportunity
Competitive compensation
Professional growth
+1
Data Engineer - Remote, Latin America
Data Engineer - Remote, Latin America

Bluelight • Belo Horizonte

Presencial
BRL 167 400 - 279 000
Competitive salary and bonuses
Generous paid-time-off policy
Technology / Office stipend
+4
Senior Spark Data Engineer - Remote Work | REF#303588
Senior Spark Data Engineer - Remote Work | REF#303588

BairesDev • Rio de Janeiro

Presencial
BRL 614 000 - 819 000
Remote work
USD compensation
Home office setup
+4
Senior Data Platform and Automation Engineer
Senior Data Platform and Automation Engineer

Lever, Inc. • Brasil

Teletrabalho
BRL 700 000 - 1 001 000
Competitive USD salary
100% remote across LATAM
Access to LATAM coworking spaces
+9