Staff Data Engineer

Emumba

Islamabad

On-site

PKR 4,200,000 - 6,500,000

Full time

9 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Emumba in Islamabad is hiring a hands-on Principal Data Engineer to architect and build large-scale data systems. This technical IC role emphasizes delivering robust data pipelines and scalable architectures.

You will design blueprints, write production code, optimize ingestion and storage layers, and mentor engineers. Collaborating with AI and backend teams, you’ll ensure data is reliable, searchable, and ready for analytics across Emumba's platforms.

Qualifications

  • 6+ years of hands-on experience designing and building production data systems.
  • Strong coding in Python and proficiency in SQL.
  • Experience with data modeling, ETL/ELT, and data lifecycle management.
  • Familiarity with event-driven systems (Kafka, Kinesis, or similar).
  • Practical experience with cloud-native data workflows (AWS S3, Lambda, Glue, or similar).
  • Solid understanding of structured, semi-structured, and unstructured data formats.

Responsibilities

  • Design and code robust data pipelines for batch and real-time use cases.
  • Define data models, schemas, and evolution strategies for scalable systems.
  • Implement reliable ingestion, transformation, and storage layers across cloud services.
  • Work hands-on with event-driven architectures, microservices, and data APIs.
  • Optimize data quality, performance, and observability.
  • Partner with AI and backend teams to make data discoverable and usable in production.
  • Contribute directly to architecture, code reviews, and performance tuning.

Skills

Python
SQL
Data modeling
ETL/ELT
Event-driven architectures
Cloud data pipelines
Data governance

Tools

Kafka
Kinesis
AWS Glue
AWS Lambda
S3
Delta Lake
Iceberg
Hudi

Job description

Staff Data Engineer

Department: Backend

Employment Type: Full Time

Location: Islamabad, Pak

Description

We’re hiring a hands-on Principal Data Engineer who can both architect and build large-scale data systems.

You’ll design the blueprints for modern data pipelines, then write the code to bring them to life. From event-driven ingestion to analytics-ready datasets, your work will power the data backbone behind Emumba’s AI and product platforms.

This is a technical IC position, not a managerial one. You’ll spend most of your time writing code, optimizing data pipelines, and mentoring engineers through real-world implementation.

Key Responsibilities
  • Design and code robust data pipelines for batch and real-time use cases.
  • Define data models, schemas, and evolution strategies for scalable systems.
  • Implement reliable ingestion, transformation, and storage layers across cloud services.
  • Work hands-on with event-driven architectures, microservices, and data APIs.
  • Optimize data quality, performance, and observability.
  • Partner with AI and backend teams to make data discoverable and usable in production.

Contribute directly to architecture, code reviews, and performance tuning.

Skills, Knowledge and Expertise
  • 6+ years of hands-on experience in designing and building production data systems.
  • Strong coding in Python and proficiency in SQL.
  • Experience with data modeling, ETL/ELT, and data lifecycle management.
  • Familiarity with event-driven systems (Kafka, Kinesis, or similar).
  • Practical experience with cloud-native data workflows (AWS S3, Lambda, Glue, or similar).
  • Solid understanding of structured, semi-structured, and unstructured data formats.
  • Experience working with vector stores or retrieval-based pipelines.
  • Familiarity with Lakehouse concepts (Delta/Iceberg/Hudi — or similar patterns).
  • Understanding of data governance, cataloging, and lineage tracking.
  • Exposure to ML data preparation or feature store design.
  • Hands-on builder attitude: owns delivery from architecture to deployment.
  • Collaborates across AI, backend, and DevOps teams.
  • Provides mentorship through examples, not supervision.

Communicates clearly and pragmatically about technical trade-offs.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Principal Data Engineer — Real-Time Data Pipelines
Principal Data Engineer — Real-Time Data Pipelines

Emumba • Islamabad

On-site
PKR 4,200,000 - 6,500,000
Staff Backend Engineer
Staff Backend Engineer

Emumba Inc. • Islamabad

On-site
PKR 1,200,000 - 1,800,000
Engineering Manager / Tech Lead – Data, AI & Development
Engineering Manager / Tech Lead – Data, AI & Development

Digifloat • Islamabad

On-site
PKR 300,000 - 600,000
Provident Fund
Annual Increment
Fully sponsored Certifications
+1
Data Engineer - Islamabad
Data Engineer - Islamabad

Yeah! Global • Islamabad

On-site
PKR 1,116,000 - 1,674,000
Lead Data Engineer (PySpark, AWS Redshift, Glue, EMR) - Hybrid Job ID: 379477
Lead Data Engineer (PySpark, AWS Redshift, Glue, EMR) - Hybrid Job ID: 379477

HireOn LLC. • Pakistan

On-site
PKR 300,000 - 800,000
Hybrid work model
Principal Data Engineer – AWS Redshift & Data Platforms Job ID: 349866
Principal Data Engineer – AWS Redshift & Data Platforms Job ID: 349866

HireOn LLC. • Lahore

On-site
PKR 2,232,000 - 3,348,000
Opportunity to work with a globally recognized client
Exposure to modern cloud data technologies
Career growth into data architecture roles
Senior Data Engineer
Senior Data Engineer

NorthBay Solutions • Lahore

On-site
PKR 1,800,000 - 2,500,000
Senior Data Engineer
Senior Data Engineer

NorthBay Solutions • Islamabad

On-site
PKR 2,500,000 - 3,500,000
Senior Data Engineer
Senior Data Engineer

NorthBay Solutions LLC • Lahore

On-site
PKR 1,500,000 - 2,000,000
Senior Data Engineer
Senior Data Engineer

NorthBay Solutions • Karachi Division

On-site
PKR 1,500,000 - 2,000,000