Staff Data Engineer

Emumba

Islamabad

On-site

PKR 2,000,000 - 4,000,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Emumba is seeking a hands-on Staff Data Engineer in Islamabad to architect and build large-scale data systems. You will design blueprints for modern data pipelines and write the code to bring them to life, powering Emumba’s AI and product platforms.

This is a technical IC position focused on coding, optimizing pipelines, and mentoring engineers through real-world implementation. You will work on event-driven ingestion, data modeling, and end-to-end data flows across cloud services, collaborating

Qualifications

  • 6+ years of hands-on experience designing and building production data systems.
  • Hands-on coding in Python and proficient SQL.
  • Experience with data modeling, ETL/ELT, and data lifecycle management.
  • Familiarity with event-driven systems (Kafka, Kinesis).
  • Experience with cloud-native data workflows (AWS S3, Lambda, Glue).
  • Understanding of various data formats (structured, semi-structured, unstructured).

Responsibilities

  • Design and code robust data pipelines for batch and real-time use cases.
  • Define data models, schemas, and evolution strategies for scalable systems.
  • Implement ingestion, transformation, and storage layers across cloud services.
  • Work hands-on with event-driven architectures, microservices, and data APIs.
  • Optimize data quality, performance, and observability.
  • Collaborate with AI and backend teams to make data discoverable in production.
  • Contribute to architecture, code reviews, and performance tuning.

Skills

Python
SQL
Data modeling
ETL/ELT
Kafka/Kinesis
Cloud data workflows
Vector stores

Job description

Staff Data Engineer

Department: Backend

Employment Type: Full Time

Location: Islamabad, Pak

Description

We’re hiring a hands-on Principal Data Engineer who can both architect and build large-scale data systems.

You’ll design the blueprints for modern data pipelines, then write the code to bring them to life. From event-driven ingestion to analytics-ready datasets, your work will power the data backbone behind Emumba’s AI and product platforms.

This is a technical IC position, not a managerial one. You’ll spend most of your time writing code, optimizing data pipelines, and mentoring engineers through real-world implementation.

Key Responsibilities
  • Design and code robust data pipelines for batch and real-time use cases.
  • Define data models, schemas, and evolution strategies for scalable systems.
  • Implement reliable ingestion, transformation, and storage layers across cloud services.
  • Work hands-on with event-driven architectures, microservices, and data APIs.
  • Optimize data quality, performance, and observability.
  • Partner with AI and backend teams to make data discoverable and usable in production.

Contribute directly to architecture, code reviews, and performance tuning.

Skills, Knowledge and Expertise
  • 6+ years of hands-on experience in designing and building production data systems.
  • Strong coding in Python and proficiency in SQL.
  • Experience with data modeling, ETL/ELT, and data lifecycle management.
  • Familiarity with event-driven systems (Kafka, Kinesis, or similar).
  • Practical experience with cloud-native data workflows (AWS S3, Lambda, Glue, or similar).
  • Solid understanding of structured, semi-structured, and unstructured data formats.
  • Experience working with vector stores or retrieval-based pipelines.
  • Familiarity with Lakehouse concepts (Delta/Iceberg/Hudi — or similar patterns).
  • Understanding of data governance, cataloging, and lineage tracking.
  • Exposure to ML data preparation or feature store design.
  • Hands-on builder attitude: owns delivery from architecture to deployment.
  • Collaborates across AI, backend, and DevOps teams.
  • Provides mentorship through examples, not supervision.

Communicates clearly and pragmatically about technical trade-offs.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Hands-On Principal Data Engineer: Data Pipelines & AI
Hands-On Principal Data Engineer: Data Pipelines & AI

Emumba • Islamabad

On-site
PKR 2,000,000 - 4,000,000
Staff Backend Engineer
Staff Backend Engineer

Emumba Inc. • Islamabad

On-site
PKR 1,200,000 - 1,800,000
Principal Data Engineer
Principal Data Engineer

Emumba • Islamabad

On-site
PKR 2,000,000 - 5,000,000
Data Engineer - Islamabad
Data Engineer - Islamabad

Yeah! Global • Islamabad

On-site
PKR 1,116,000 - 1,674,000
Hands-on Principal Data Engineer: Architect & Build Pipelines
Hands-on Principal Data Engineer: Architect & Build Pipelines

Emumba • Islamabad

On-site
PKR 2,000,000 - 5,000,000
Engineering Manager / Tech Lead – Data, AI & Development
Engineering Manager / Tech Lead – Data, AI & Development

Digifloat • Islamabad

On-site
PKR 300,000 - 600,000
Provident Fund
Annual Increment
Fully sponsored Certifications
+1
Lead Data Engineer (PySpark, AWS Redshift, Glue, EMR) - Hybrid Job ID: 379477
Lead Data Engineer (PySpark, AWS Redshift, Glue, EMR) - Hybrid Job ID: 379477

HireOn LLC. • Pakistan

On-site
PKR 300,000 - 800,000
Hybrid work model
Principal Data Engineer – AWS Redshift & Data Platforms Job ID: 349866
Principal Data Engineer – AWS Redshift & Data Platforms Job ID: 349866

HireOn LLC. • Lahore

On-site
PKR 2,232,000 - 3,348,000
Opportunity to work with a globally recognized client
Exposure to modern cloud data technologies
Career growth into data architecture roles
Principal Data Engineer – AWS Redshift & Data Platforms
Principal Data Engineer – AWS Redshift & Data Platforms

HireOn • Karachi Division

On-site
PKR 1,674,000 - 2,790,000
Hybrid work flexibility
Exposure to modern technologies
Career growth opportunities
Senior Data Engineer
Senior Data Engineer

NorthBay Solutions LLC • Lahore

On-site
PKR 1,500,000 - 2,000,000