Senior Data Engineer

Convo

Islamabad

On-site

PKR 2,000,000 - 4,500,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Convo is seeking a Senior Data Engineer to build reliable data pipelines and connectors powering its CPG-focused agentic platform. You’ll develop scalable batch and streaming pipelines using Python, Spark, Databricks, Kafka, dbt, and Delta Lake, transforming source data into governed RAW and CUBE layers while ensuring data quality, performance, observability, and reliability.

You will also build data-product and feature-serving interfaces that enable ML models, agents, and downstream

Qualifications

  • Minimum experience: 5+ years in data engineering, including 3+ years delivering production Spark or Python pipelines.
  • Advanced Python and production Apache Spark/Databricks development.
  • Kafka or comparable event-streaming technology.
  • Strong SQL performance tuning and dimensional/data-product implementation.
  • dbt and Delta Lake, including incremental patterns and table optimization.
  • Batch and streaming reliability patterns: idempotency, checkpointing, replay and late-arriving data.
  • Automated data quality, observability and schema-contract testing.

Responsibilities

  • Develop Python/Spark ingestion and transformation pipelines from source systems into RAW and Enriched CUBE layers.
  • Implement streaming paths alongside scheduled workloads.
  • Create dbt/SQL transformations, data-quality controls and contract-validation checks.
  • Optimize Delta Lake layout, SQL performance, partitioning and incremental processing.
  • Build feature-serving and data-product interfaces for models, agents and applications.
  • Instrument pipelines for lineage, freshness, throughput, failure recovery and cost.

Skills

Python
Apache Spark
Databricks
Kafka
dbt
Delta Lake
SQL tuning
Data quality
Observability
Contract-validation

Job description

Job Summary

CONVO is seeking a Senior Data Engineer to build reliable data pipelines and connectors powering its CPG-focused agentic platform. You’ll develop scalable batch and streaming pipelines using Python, Spark, Databricks, Kafka, dbt, and Delta Lake, transforming source data into governed RAW and CUBE layers while ensuring data quality, performance, observability, and reliability. You’ll also build data-product and feature-serving interfaces that enable ML models, agents, and downstream applications to consume trusted data efficiently.

Technical mission
  • Build reliable connectors and RAW-to-CUBE pipelines across batch and streaming paths, and expose governed data and ML features to downstream services.
Key responsibilities
  • Develop Python/Spark ingestion and transformation pipelines from source systems into RAW and Enriched CUBE layers.
  • Implement streaming paths alongside scheduled workloads.
  • Create dbt/SQL transformations, data-quality controls and contract-validation checks.
  • Optimize Delta Lake layout, SQL performance, partitioning and incremental processing.
  • Build feature-serving and data-product interfaces for models, agents and applications.
  • Instrument pipelines for lineage, freshness, throughput, failure recovery and cost.
Required technical capabilities
  • Minimum experience: 5+ years in data engineering, including 3+ years delivering production Spark or python pipelines.
  • Advanced Python and production Apache Spark/Databricks development.
  • Kafka or comparable event-streaming technology.
  • Strong SQL performance tuning and dimensional/data-product implementation.
  • dbt and Delta Lake, including incremental patterns and table optimization.
  • Batch and streaming reliability patterns: idempotency, checkpointing, replay and late-arriving data.
  • Automated data quality, observability and schema-contract testing.
Preferred experience
  • ML feature stores or online/offline feature consistency.
  • CPG, retail, ERP, POS or syndicated-data pipelines.
  • Kubernetes-based data workloads and cloud cost optimization.
Expected deliverables / acceptance evidence
  • Production connectors and RAW-to-CUBE pipelines with automated tests.
  • Batch/streaming operational dashboards and recovery procedures.
  • Documented data products and feature-serving interfaces.
  • Performance and cost baselines for the implemented workloads.
Primary interfaces
  • Works under the Data Architect with source-system owners, ML/optimization teams, platform infrastructure and downstream application teams.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer: Scalable Pipelines & ML Features
Senior Data Engineer: Scalable Pipelines & ML Features

Convo • Islamabad

On-site
PKR 2,000,000 - 4,500,000
Senior Data Engineer
Senior Data Engineer

Systems Limited • Punjab

On-site
PKR 2,400,000 - 3,600,000
Data Engineer
Data Engineer

Archisurance • Lahore

On-site
PKR 1,200,000 - 1,800,000
Senior Data Engineer (Pyspark, Databricks)
Senior Data Engineer (Pyspark, Databricks)

Strategic Systems International • Lahore

On-site
PKR 1,800,000 - 3,200,000
Associate Data Engineer Team Lead: Build Pipelines & Impact
Associate Data Engineer Team Lead: Build Pipelines & Impact

Devsinc • Lahore

On-site
Associate Team Lead- Data Engineer
Associate Team Lead- Data Engineer

Devsinc • Lahore

On-site
PKR 2,232,000 - 3,906,000
Senior Data Architect
Senior Data Architect

Convo • Islamabad

On-site
PKR 3,000,000 - 6,000,000
Medical benefits
Free lunch
Performance-based increments
+1
Data Engineer
Data Engineer

Logiciel Services, Llc • Karachi Division

On-site
PKR 600,000 - 1,200,000
Sr. Data Scientist
Sr. Data Scientist

Abacus Global • Lahore

On-site
PKR 3,000,000 - 4,500,000
Senior Data Engineer — Real-Time Data Platform Lead
Senior Data Engineer — Real-Time Data Platform Lead

Systems Limited • Punjab

On-site
PKR 2,400,000 - 3,600,000