Staff Engineer - Data Engineer

Nagarro

Rio de Janeiro

Presencial

BRL 180 000 - 320 000

Tempo integral

Há 5 dias
Torna-te num dos primeiros candidatos

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Nagarro is seeking a Senior Data Architect to bridge data modeling and KM domain expertise, translating unstructured content into secure data products within Databricks Unity Catalog. You will define data products, metadata standards, and governance across knowledge sources.

Responsibilities include collaborating with data engineers and stakeholders, establishing repeatable modeling standards, and guiding privacy classifications to support downstream discovery tools used by Sage/Glean-based AI

Qualificações

  • 5+ years of experience in data modeling, data architecture, or information architecture.
  • Exposure to unstructured or semi-structured data, not purely relational modeling.
  • Experience in KM, content management, or enterprise search domain.
  • Hands-on with a modern data catalog; Databricks Unity Catalog preferred.
  • Ability to define data domains and data product boundaries in a large org.
  • Knowledge of metadata management and taxonomy design.
  • Understanding data security and access control in a lakehouse environment.
  • Collaborates with data engineering and governance stakeholders to align data products.

Responsabilidades

  • Design data models for unstructured content from KM pipelines and artifacts.
  • Define data product boundaries and ownership for reusable assets.
  • Establish metadata standards and tagging taxonomies for classification.
  • Apply security and sensitivity classifications per governance and privacy rules.
  • Register and maintain data products in Unity Catalog with schemas and lineage.
  • Partner with data engineers to align ingestion and storage patterns with the model.
  • Collaborate with Knowledge/Research/Architecture stakeholders for downstream use.
  • Support privacy and legal reviews with clear documentation and classifications.
  • Create repeatable modeling standards/playbooks for future KM scaling.

Conhecimentos

Data modeling
Data architecture
Information architecture
Communication skills
Databricks Unity Catalog

Ferramentas

Databricks Unity Catalog

Descrição da oferta de emprego

We are a Digital Product Engineering company that is scaling in a big way! We build products, services, and experiences that inspire, excite, and delight. We work at scale — across all devices and digital mediums, and our people exist everywhere in the world (15000+ experts across 26 countries, to be exact). Our work culture is dynamic and non-hierarchical. We are looking for great new colleagues. That is where you come in!

Job Description

This role bridges data engineering discipline with KM domain expertise, translating raw unstructured content (documents, case files, informal knowledge captures, chat/email extracts, etc.) into well-defined, discoverable, and secure data products within Databricks Unity Catalog.

Key Responsibilities

  • Design logical and physical data models for unstructured and semi-structured content (documents, case artifacts, K-Slices, extracted knowledge fragments, metadata records) originating from KM pipelines such as case mining and informal knowledge capture workflows.
  • Define domain boundaries and ownership for data products — determining what constitutes a discrete, reusable data product versus a raw or intermediate asset.
  • Establish metadata standards and tagging taxonomies (content type, practice/domain, provenance, confidentiality, freshness, lineage) to ensure consistent classification across knowledge sources.
  • Assign and enforce security and sensitivity classifications on data products in line with firm data governance, privacy, and legal/risk requirements.
  • Register, document, and maintain data products in Databricks Unity Catalog, including schemas, access grants, lineage, and catalog-level metadata.
  • Partner with data engineers building Databricks pipelines to ensure ingestion, transformation, and storage patterns align to the modeled domain structure.
  • Collaborate with Knowledge Products, Research Products, and Architecture/Data/Technology stakeholders to align data product design with downstream consumption needs (e.g., surfacing in Sage/Glean, AI agent retrieval).
  • Support privacy and legal review processes by ensuring data products are classified and documented to enable timely sign-off.
  • Establish and document repeatable modeling standards/playbooks so future data products can be onboarded consistently as the KM platform scales.

Required Qualifications

  • 5+ years of experience in data modeling, data architecture, or information architecture, with meaningful exposure to unstructured or semi-structured data (not purely relational/transactional modeling).
  • Direct experience working in or adjacent to Knowledge Management, content management, or enterprise search domain — understands how documents, case files, or knowledge artifacts differ from standard transactional data.
  • Hands-on experience with a modern data catalog; Databricks Unity Catalog experience strongly preferred.
  • Demonstrated ability to define data domains and data product boundaries in a large, multi-stakeholder organization.
  • Practical knowledge of metadata management: tagging schemas, taxonomies, controlled vocabularies, or ontology design.
  • Understanding of data security/sensitivity classification frameworks and how they map to access control in a lakehouse environment.
  • Experience partnering with data engineering teams on ingestion and pipeline design (not required to write production pipeline code, but must speak the language).
  • Strong written and verbal communication skills; able to translate technical modeling decisions into business-readable rationale for KM stakeholders and governance reviewers.

Preferred Qualifications

  • Experience with enterprise knowledge platforms (e.g., Glean, SharePoint, ServiceNow) or AI-powered retrieval systems.
  • Familiarity with Databricks Delta Lake, Delta Sharing, or Lakehouse Federation.
  • Prior experience in professional services, consulting, or a similar document/case-intensive knowledge environment.
  • Exposure to Legal/Risk/Privacy review processes for data classification and access approvals.
  • Background in library science, information science, or applied ontology is a plus but not required. Success Metrics (First 6–12 Months)
  • Domain model and metadata taxonomy defined and adopted for at least one major KM data product line (e.g., case mining outputs, informal knowledge K-Slices).
  • Data products registered and discoverable in Unity Catalog with correct security classifications applied.
  • Documented, repeatable modeling standard that engineering and future modelers can apply without re-litigating domain boundaries each time.
  • Reduced turnaround time on privacy/legal classification reviews due to upfront, consistent metadata and tagging.
Qualifications

Must have skills: Data Modeling (Strong), Databricks

Good to have skills: BI Schema Design - General Experience

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Staff Engineer - Data Engineer
Staff Engineer - Data Engineer

Nagarro • Brasil

Presencial
BRL 180 000 - 300 000
Director Associate Distinguished Engineer - Solution Architect
Director Associate Distinguished Engineer - Solution Architect

Nagarro • Rio de Janeiro

Híbrido
BRL 300 000 - 600 000
Associate Distinguished Engineer - Solution Architect
Associate Distinguished Engineer - Solution Architect

Nagarro • Rio de Janeiro

Presencial
BRL 250 000 - 420 000
Analytics Engineer – Data Modeling
Analytics Engineer – Data Modeling

Jobtailor • Santana de Parnaíba

Presencial
BRL 180 000 - 280 000
Data Engineer – Databricks
Data Engineer – Databricks

Jobtailor • São Paulo

Presencial
BRL 180 000 - 280 000
Sr. Solutions Engineer
Sr. Solutions Engineer

Cacheflow • São Paulo

Presencial
BRL 609 000 - 915 000
Sr. Solutions Engineer
Sr. Solutions Engineer

WinsAbove • São Paulo

Presencial
BRL 300 000 - 540 000
Sr. Solutions Engineer
Sr. Solutions Engineer

Databricks • São Paulo

Presencial
BRL 180 000 - 260 000
INTL LATAM Data Engineer
INTL LATAM Data Engineer

Insight Global • Bezerros

Presencial
BRL 120 000 - 180 000
Solutions Architect - Lakebase
Solutions Architect - Lakebase

WinsAbove • São Paulo

Presencial
BRL 180 000 - 240 000