DataOps Engineer

Mars

Arujá

Presencial

BRL 120 000 - 210 000

Tempo integral

Há 3 dias
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Transforma esta função numa entrevista — um currículo e uma carta de apresentação criados à volta do que este empregador procura.

Ultrapassa os filtros ATS

Resumo da oferta

Mars Data Platform Services (DPS) is modernizing its enterprise data platform on Azure Databricks and Unity Catalog. We are building a metadata-driven ingestion engine to automate data onboarding, schema evolution, and pipeline configuration.

The DataOps Engineer will design scalable landing zones, automate Unity Catalog assets, and drive CI/CD practices with dashboards and playbooks to keep 24/7 data pipelines reliable.

Qualificações

  • 4+ years of hands-on data engineering experience building reliable batch and streaming data pipelines using Python/PySpark, SQL, and cloud data warehouses or lakehouses (Azure Databricks / Delta Lake).
  • Metadata-driven/config-driven frameworks with central control tables.
  • Programmatic Unity Catalog & Landing Zone architecture using Databricks APIs and cloud storage.
  • DataOps & automation: CI/CD, version control, automated testing, and modular packaging (Databricks Asset Bundles or Terraform).

Responsabilidades

  • Build & Scale the Metadata-Driven Ingestion Engine with configurable onboarding of new data sources.
  • Lead Landing Zone Reliability and Data Quality safeguards across cloud storage and Unity Catalog.
  • Automate Unity Catalog assets and long-term data archiving with policy-driven retention.
  • Champion DataOps standards by enabling automated testing, deployment, and monitoring for pipelines.

Conhecimentos

Python/PySpark
SQL
Azure Databricks/Delta Lake
Metadata-driven frameworks
DataOps & CI/CD

Ferramentas

Databricks
Terraform
Git

Descrição da oferta de emprego

Job Description:

At Mars Data Platform Services (DPS), we are transforming our enterprise data platform across Azure Databricks and Unity Catalog into an automated, self-service product. Instead of hand-building custom pipelines for every new dataset, we are engineering an intelligent, metadata-driven ingestion engine that configures and scales pipelines automatically.

We are looking for a DataOps Engineer to lead the technical evolution of our data inflow and landing zone ecosystem. In this role, you will build the central framework that turns pipeline creation, data onboarding, storage provisioning, and data archiving into automated, configuration-driven processes.

What Will Be Your Key Responsibilities?
  • Build & Scale the Metadata-Driven Ingestion Engine: Design and enhance our central metadata framework so onboarding new data sources requires simple configuration updates rather than writing new pipeline code. Ensure the engine dynamically orchestrates data ingestion, schema evolution, and merge logic reliably across the platform.
  • Lead Landing Zone Reliability & Data Quality Safeguards: Design and maintain high-throughput landing zones across cloud storage and Unity Catalog. Build automated safeguards that isolate schema errors or corrupted data into quarantine areas and dead-letter queues before they can impact production data tables.
  • Automate Unity Catalog Assets & Long-Term Archiving: Automate the programmatic creation and setup of storage locations, schemas, and tables in Unity Catalog. Implement automated lifecycle policies that handle time-based data retention and migrate older data into low-cost archival storage seamlessly.
  • Champion DataOps Standards & Operational Enablement: Drive modern software engineering standards into our data workflows through automated CI/CD testing and deployment. Create reusable deployment templates, monitoring dashboards, and clear operational playbooks that empower our 24/7 support teams to monitor and maintain pipelines with confidence.
What Are We Looking For?
Essential Requirements:
  • Core Data Engineering Background: 4+ years of hands-on data engineering experience building reliable batch and streaming data pipelines using Python/PySpark, SQL, and cloud data warehouses or lakehouses (Azure Databricks / Delta Lake).
  • Metadata-Driven Mindset: Proven track record designing or maintaining config-driven/metadata-driven frameworks where pipeline generation, schema creation, Delta merge logic, and scheduling are driven dynamically by central control tables rather than hand-coded per dataset.
  • Programmatic Unity Catalog & Landing Zone Architecture: Practical expertise managing and provisioning Databricks Unity Catalog storage assets (Volumes, External Locations, Managed/External Tables) programmatically via code/APIs, alongside high-throughput Azure Data Lake Storage (ADLS Gen2) landing zones.
  • DataOps & Automation Practices: Experience treating data pipelines like software products—applying automated CI/CD deployment, version control (Git), automated testing, and modular packaging (such as Databricks Asset Bundles or Terraform).
Nice-to-Haves:
  • Experience with declarative pipeline tools (such as Delta Live Tables or dbt).
  • Practical experience designing automated data retention schedules and long-term cloud archival storage.
  • Familiarity with supporting or partnering with operational support teams and Managed Service Providers (MSPs

#TBdigital

Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

DataOps Engineer
DataOps Engineer

Mars, Incorporated and its Affiliates • Guararema

Presencial
BRL 180 000 - 300 000
Data Platform Lead / Cloud Systems Architect
Data Platform Lead / Cloud Systems Architect

GyanSys Inc. • Brasil

Presencial
BRL 320 000 - 520 000
Specialist Associate Principal Engineer - DataOps
Specialist Associate Principal Engineer - DataOps

SmartRecruiters, Inc. • Rio de Janeiro

Híbrido
BRL 120 000 - 180 000
Associate Principal Engineer - DataOps
Associate Principal Engineer - DataOps

SmartRecruiters, Inc. • Rio de Janeiro

Presencial
BRL 180 000 - 260 000
Senior Data Engineer
Senior Data Engineer

Quartile • Brasil

Presencial
BRL 180 000 - 320 000
Data Engineer
Data Engineer

Pride Global • São Paulo

Presencial
BRL 167 400 - 279 000
Data Platform Engineer Id90126
Data Platform Engineer Id90126

Agileengine • Curitiba

Presencial
BRL 180 000 - 280 000
Professional growth
Competitive USD-based compensation
Mentorship and TechTalks
Data Platform Engineer Id90126
Data Platform Engineer Id90126

Agileengine • Campinas

Presencial
BRL 470 000 - 731 000
Professional growth
Competitive USD-based compensation
Mentorship and TechTalks
Data Platform Engineer Id90126
Data Platform Engineer Id90126

Agileengine • São Bernardo do Campo

Presencial
BRL 470 000 - 783 000
Professional growth
Competitive USD-based compensation
Data Platform Engineer Id90126
Data Platform Engineer Id90126

Agileengine • Brasília

Presencial
BRL 627 000 - 940 000
Professional growth
Education budget
Team activities