Senior Data Engineer

careers-quartile

Brasil

Presencial

BRL 180 000 - 300 000

Tempo integral

Há 11 dias

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Quartile is seeking a Senior Data Engineer to own ingestion, modeling, and serving layers in a Databricks-based lakehouse. You will implement medallion architectures, CDC patterns, and scalable data contracts while collaborating with AI engineers and product teams. This is a PJ contract in Brazil with focus on reliability, cost, and performance.

You will work with Python, SQL, and Azure tools, ensuring data quality, governance, and end-to-end ownership from data ingestion to AI-ready outputs.

Qualificações

  • Strong software engineering fundamentals in Python and advanced SQL
  • Deep hands-on experience with Databricks: Spark/PySpark, Delta Lake, Unity Catalog, Workflows, and SQL warehouses
  • Proven experience designing data pipelines and data models in production
  • Experience integrating third-party APIs at scale, with auth, pagination, throttling, and backfills
  • Familiarity with Azure cloud services and IaC (Terraform)
  • Data quality through testing, monitoring, and alerting

Responsabilidades

  • Design, build, and operate ingestion pipelines from advertising and business platforms via REST APIs and event streams
  • Model curated data layers in Delta Lake under Unity Catalog with clear contracts and governance
  • Build the serving layer for AI tools and agents with low latency data stores
  • Own data quality and health monitoring as code with alerting and triage paths
  • Deliver infrastructure through Terraform and CI/CD with proper environment promotion
  • Collaborate with AI engineers, software engineers, and product teams to meet contracts
  • Instrument observability using OpenTelemetry and structured logging
  • Document and maintain the codebase following best practices

Conhecimentos

Python
SQL
Problem-solving
Autonomy
Analytical thinking

Formação académica

Bachelor's degree or higher in Computer Science/IT

Ferramentas

Databricks
Delta Lake
Unity Catalog
Terraform
CI/CD
OpenTelemetry

Descrição da oferta de emprego

WHO WE ARE:

Quartile, the world's largest retail media optimization platform, is a trusted partner for multichannel e-commerce success. Through unmatchedexpertiseand patented AI technology, we fuel growth for 5,300+ brands and sellers worldwide and manage an annual ad spend exceeding$2 billion. The award-winning platform covers major marketplaces and ad channels foroptimalreach. The result is unprecedented granularity, smarter budgeting, and bespoke solutions for retailers.

Quartile is proud to be an equal opportunity employer with employees stemming from a wide range of backgrounds and experiences. As a business, we value the enrichment that diversity brings to our organization and are committed to a culture that creates a sense of inclusion and belonging. We welcome new perspectives and affirm that all employment decisions are made without regard to race, color, ancestry, religion, national origin, age, familial or marital status, sex, sexual orientation, pregnancy, gender identity or expression, disability, genetic information, veteran status, or any other classification protected by federal, state, or local law.

About Sciene

At Sciene, the mission is to empower professional services firms with cutting-edge, customized AI solutions — enhancing automation, analytics, and optimization across industries while prioritizing security, cost efficiency, and state-of-the-art technology. Our flagship product, the Sciene AI Companion, is an autonomous customer success platform deployed across Quartile — the world's largest retail media optimization platform, managing performance marketing for 1,000+ brands. It automates relationship-heavy enterprise workflows end to end: generating personalized email replies in the CSM's own voice (8x faster), building full presentation decks for client meetings (12x faster), and detecting and diagnosing account fluctuations before anyone has to ask (6x faster). None of this replaces human judgment — it removes the work that was getting in the way of it. Read more about how we built it: Sciene AI Companion: Building an Autonomous Customer Success Platform on Databricks

OVERVIEW:

Every answer the AI Companion gives is only as good as the data underneath it. Sciene runs a production lakehouse on Databricks and Azure that ingests advertising, CRM, and conversational data from a dozen upstream systems, curates it under Unity Catalog governance, and serves it back to autonomous agents as tools — not as dashboards.

The Senior Data Engineer owns that substrate end to end: ingestion from messy third-party APIs, modeling and governance in Unity Catalog, the serving layer agents query at runtime, and the health monitoring that tells us something broke before a CSM finds out in a client meeting. Infrastructure is code, deploys go through CI/CD, and quality is measured — not assumed. You write the Terraform, you own the alerts, you get paged by your own monitors.

This is also a role where you work with AI, not just for it. Our engineers ship with coding agents (Claude Code and our MCP-integrated toolchain) as part of the normal workflow, and your consumer is an autonomous agent making decisions on top of what you produce — which raises the bar on correctness, freshness, and how the data is shaped.

REQUIREMENTS:
  • Strong software engineering fundamentals in Python and advanced SQL — this role builds tested, deployed, version-controlled pipelines, not one-off notebooks
  • Deep hands-on experience with Databricks: Spark/PySpark, Delta Lake, Unity Catalog, Workflows, and SQL warehouses, including performance tuning of real workloads
  • Proven experience designing data pipelines and data models in production: medallion or equivalent layered architectures, incremental processing, CDC, and dimensional modeling
  • Experience integrating third-party APIs at scale, with the operational maturity that implies (auth, pagination, throttling, partial failure, backfills, idempotency, schema evolution)
  • Familiarity with cloud platforms (we run on Azure — ADLS, Container Apps, Key Vault, API Management, Entra ID)
  • Experience with Infrastructure as Code (Terraform) and CI/CD pipelines, plus solid git and code-review habits
  • A working definition of data quality: tests, expectations, monitoring, and the discipline to catch problems upstream rather than explaining them downstream
  • Practical experience using AI coding assistants/agents in real engineering work, and the critical eye to know when their output is wrong
  • Excellent problem-solving and analytical skills, and the autonomy expected of a senior engineer: you own a problem end to end — from framing to shipped, monitored outcome — and are accountable for the result, not just the merge
  • Bachelor's degree or higher in Computer Science, Information Technology, or a related field — or equivalent practical experience
PREFERRED QUALIFICATIONS:
  • Experience serving data to LLM-powered applications, including retrieval and context design
  • Experience with operational stores alongside the lakehouse (Databricks Lakebase, PostgreSQL, MongoDB)
  • Observability stacks: OpenTelemetry, Grafana, Loki, structured logging
  • Delta Sharing for secure data exchange with clients and partners
  • Domain background in retail media, digital advertising, or marketplace data
  • Databricks or Azure certifications
WHAT YOU'LL DO:
  • Design, build, and operate ingestion pipelines from advertising and business platforms (Amazon Ads, Google Ads, Walmart Connect, Criteo, Salesforce, and internal services) via REST APIs, Delta Sharing, object storage, and event streams
  • Model curated data layers in Delta Lake under Unity Catalog: medallion architecture, incremental and CDC patterns, clear contracts between layers, lineage, and row/column-level governance
  • Build the serving layer for AI: SQL views, Unity Catalog functions, and low-latency stores exposed to agents as tools through the Model Context Protocol (MCP) behind Azure API Management — designing for what an agent needs to reason well, not just for what a BI tool needs to render
  • Own data quality and health monitoring as code: freshness, volume, schema-drift and business-rule expectations, with alerting wired into Grafana and a clear triage path when something fires
  • Ship everything through Terraform and CI/CD — Databricks and Azure resources, jobs, permissions, and environment promotion from dev to prod, with no manual clicking in the console
  • Own performance and cost: warehouse and cluster sizing, partitioning and liquid clustering, job orchestration in Databricks Workflows / Lakeflow, and continuous scrutiny of what each pipeline actually costs to run
  • Work fluently with AI coding agents as part of your daily engineering practice — writing precise specs, keeping repository context and documentation in a state where agents produce good output, and reviewing what they generate with real judgment
  • Collaborate with cross-functional teams — AI engineers, software engineers, and product — turning product requirements into data the platform can actually serve, and treating internal consumers as customers with contracts and expectations
  • Instrument what you ship with OpenTelemetry and structured logging, and improve reliability and latency in production
  • Document and maintain the codebase, ensuring code quality and adherence to best practices

*This is a PJ contract based in Brazil.

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior AI Engineer
Senior AI Engineer

Quartile • Brasil

Presencial
BRL 180 000 - 240 000
Data Engineer
Data Engineer

Pride Global • São Paulo

Presencial
Senior Analytics Engineer (Databricks, Spark)
Senior Analytics Engineer (Databricks, Spark)

exadelinc • São Paulo

Presencial
BRL 300 000 - 420 000
1104 | Senior Data (Databricks) Engineer
1104 | Senior Data (Databricks) Engineer

Intetics • Brasil

Presencial
BRL 180 000 - 300 000
Lead Data Engineer ID71008
Lead Data Engineer ID71008

AgileEngine, LLC. • Salvador

Presencial
BRL 260 000 - 420 000
Growth without limits
Competitive compensation
Flexibility: 100% remote
+3
Lead Data Engineer ID71008
Lead Data Engineer ID71008

AgileEngine, LLC. • Brasília

Presencial
BRL 300 000 - 600 000
Growth opportunities
Competitive pay
Remote first
+3
Senior Data Engineer
Senior Data Engineer

Perform • São Paulo

Presencial
BRL 200 000 - 250 000
Data Engineer, AI & Analytics
Data Engineer, AI & Analytics

Powerdigitalmarketing • Brasil

Presencial
BRL 180 000 - 280 000
Senior Data Engineer ID81743
Senior Data Engineer ID81743

AgileEngine, LLC. • Brasil

Presencial
BRL 180 000 - 280 000
Growth without limits
Competitive compensation
Flexibility: 100% remote
Senior Data Engineer ID81743
Senior Data Engineer ID81743

AgileEngine, LLC. • Campinas

Presencial
BRL 180 000 - 300 000
Growth without limits
Competitive compensation
Remote work
+3