Analytics Engineer – Data Modeling

Jobtailor

Santana de Parnaíba

Presencial

BRL 180 000 - 280 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Jobtailor in São Paulo seeks an experienced data engineer to design and maintain the enterprise data integration layer on Databricks, implementing Data Vault patterns and ensuring data quality. You will drive standardized, traceable analytics and collaborate with BI and data teams to meet strategic data goals.

Responsibilities include managing the Silver/Gold layers, enforcing data lineage in Unity Catalog, and optimizing pipelines with governance and security in mind.

Qualificações

  • Advanced SQL with analytic queries and optimization.
  • Python for data engineering and analytics.
  • Hands-on Apache Spark experience, preferably PySpark.
  • Experience with Databricks, Delta Lakehouse, DLT, Unity Catalog.
  • Strong data modeling knowledge (Data Vault and dimensional models).
  • Familiarity with Medallion architecture (Bronze, Silver and Gold).
  • Experience with incremental ingestion and historical data handling.
  • Experience with pipeline orchestration and observability (freshness, quality and alerts).
  • Data governance and information security (masking, RBAC/ABAC, LGPD) desirable.
  • Knowledge of Git, CI/CD and data testing.

Responsabilidades

  • Design, model, implement and maintain the enterprise data integration layer, primarily the Silver layer on Databricks, using Data Vault modeling (Raw Vault and Business Vault)
  • Integrate data from multiple systems, ensuring standardization, deduplication, normalization and traceability of information
  • Apply identity resolution techniques across distinct sources to ensure correct unification of entities such as customers, contracts and transactions
  • Normalize and standardize analytical attributes, including handling of marketing identifiers and digital channels (e.g., UTM), ensuring consistency for analytical consumption
  • Build and maintain unified entities by applying structural business rules and ensuring semantic coherence across domains
  • Provide the integrated and semantic data layer as the official source for Business Intelligence consumption, supporting dimensional modeling in the Gold layer
  • Collaborate with BI, Data Engineering and business areas to gather analytical requirements and identify information gaps
  • Ensure new analytical projects adopt the integrated data layer as the corporate standard from the start
  • Implement data quality practices, including validations, automated tests and continuous monitoring
  • Document, catalog and maintain data lineage in Unity Catalog, with complete metadata and clear ownership definitions
  • Define and maintain naming, semantic and data modeling standards together with the relevant stakeholders
  • Orchestrate, monitor and optimize pipelines under your responsibility, ensuring performance, stability, security and cost control
  • Work with the company’s Governance function to enforce access policies, information security and regulatory compliance

Conhecimentos

SQL
Python
PySpark
Data Vault modeling
Dimensional modeling
Data governance
Git/CI-CD
Data lineage
Analytics
ETL orchestration

Ferramentas

Databricks
Unity Catalog
Delta Lakehouse
DLT
Git

Descrição da oferta de emprego

Responsibilities
  • Design, model, implement and maintain the enterprise data integration layer, primarily the Silver layer on Databricks, using Data Vault modeling (Raw Vault and Business Vault)
  • Integrate data from multiple systems, ensuring standardization, deduplication, normalization and traceability of information
  • Apply identity resolution techniques across distinct sources to ensure correct unification of entities such as customers, contracts and transactions
  • Normalize and standardize analytical attributes, including handling of marketing identifiers and digital channels (e.g., UTM), ensuring consistency for analytical consumption
  • Build and maintain unified entities by applying structural business rules and ensuring semantic coherence across domains
  • Provide the integrated and semantic data layer as the official source for Business Intelligence consumption, supporting dimensional modeling in the Gold layer
  • Collaborate with BI, Data Engineering and business areas to gather analytical requirements and identify information gaps
  • Ensure new analytical projects adopt the integrated data layer as the corporate standard from the start
  • Implement data quality practices, including validations, automated tests and continuous monitoring
  • Document, catalog and maintain data lineage in Unity Catalog, with complete metadata and clear ownership definitions
  • Define and maintain naming, semantic and data modeling standards together with the relevant stakeholders
  • Orchestrate, monitor and optimize pipelines under your responsibility, ensuring performance, stability, security and cost control
  • Work with the company’s Governance function to enforce access policies, information security and regulatory compliance
Requirements
  • SQL (Advanced proficiency, including analytical queries, modeling and optimization)
  • Python (applied to data engineering and analytics)
  • Hands‑on experience with Apache Spark (preferably PySpark)
  • Experience with Databricks (Delta Lakehouse, notebooks, DLT, Unity Catalog)
  • Strong knowledge of data modeling (Data Vault and dimensional star schema)
  • Familiarity with the Medallion architecture (Bronze, Silver and Gold layers)
  • Experience with incremental ingestion and historical data handling
  • Experience with pipeline orchestration and observability (freshness, quality and alerts)
  • Desirable: experience with Data Governance and information security applied to data (masking, RBAC/ABAC and LGPD)
  • Knowledge of code versioning, Git, CI/CD and data testing
Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Data Engineer
Senior Data Engineer

Jobtailor • São Paulo

Presencial
BRL 120 000 - 180 000
Data Engineer – Databricks
Data Engineer – Databricks

Jobtailor • São Paulo

Presencial
BRL 180 000 - 280 000
Data Engineer
Data Engineer

Luxoft • Brasil

Presencial
BRL 200 000 - 420 000
AI Data Engineer
AI Data Engineer

remopt • São Paulo

Presencial
Data Engineer – Senior
Data Engineer – Senior

Jobtailor • São Paulo

Presencial
BRL 180 000 - 300 000
Senior Data Intelligence Analyst
Senior Data Intelligence Analyst

Jobtailor • Rio de Janeiro

Presencial
BRL 180 000 - 260 000
Data Engineer I – Digital and IT Governance
Data Engineer I – Digital and IT Governance

Jobtailor • Goiânia

Presencial
BRL 120 000 - 240 000
Senior Data Engineer
Senior Data Engineer

Perform • São Paulo

Presencial
BRL 200 000 - 250 000
Data Engineer, Mid-level
Data Engineer, Mid-level

Jobtailor • São Paulo

Presencial
BRL 120 000 - 200 000
Data Analyst I – BI
Data Analyst I – BI

Jobtailor • Porto Alegre

Presencial