Data Engineer

Damco Solutions

Gurugram District

Hybrid

INR 4,000,000 - 7,000,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Damco Solutions is seeking a Data Engineering Lead with strong AWS data services and retail data ecosystems to design, build, and optimize scalable pipelines that ingest MMS and POS data into a centralized ODS for analytics.

You will guide the team through data modeling, CDC upserts, and data governance, partner with microservices and DevOps to ensure real-time data availability and cost-efficient processing using AWS Glue, Redshift, S3, Kinesis, and Lambda.

Qualifications

  • 5+ years in data engineering, incl. 2+ years in a lead role.
  • Hands-on with AWS data services (Glue, S3, Redshift, EMR).
  • Proficient in PySpark/Spark for large-scale processing.
  • Experience designing ODS layers using NoSQL (Couchbase, DynamoDB, MongoDB).
  • Strong ETL/ELT design, ingestion and transformation pipelines.
  • Experience with MMS/POS retail data preferred.
  • CI/CD pipelines experience (GitLab, CodePipeline).
  • Knowledge of data governance, quality, metadata frameworks.
  • Experience supporting microservices data platforms.
  • Excellent communication and stakeholder management.

Responsibilities

  • Design and implement scalable ETL/ELT pipelines ingesting data from MMS/POS/third-party systems into AWS-based platforms.
  • Build and maintain a centralized ODS using NoSQL tech, optimized for low-latency access.
  • Develop data processing workflows with PySpark and AWS services (Glue, EMR).
  • Create denormalized, API-ready data models for downstream services.
  • Implement idempotent processing, CDC upsert strategies, and data reconciliation.
  • Leverage S3, Redshift, or Aurora for storage and querying of structured data.
  • Implement batch and streaming ingestion patterns (Kinesis, Lambda).
  • Apply performance tuning to improve pipeline efficiency and cost-effectiveness.
  • Define data governance, quality, and metadata standards across the platform.
  • Collaborate with DevOps to maintain CI/CD pipelines (CodePipeline, GitLab).
  • Provide technical leadership, peer reviews, and mentorship to the data team.
  • Work with microservices and application teams for seamless ODS integration via APIs/events.

Skills

PySpark
ETL/ELT design
Data pipelines
NoSQL (Couchbase)
AWS Glue
AWS S3
Redshift
CDC upsert
API/data modeling
CI/CD
Mentorship
Stakeholder management

Tools

GitLab
AWS CodePipeline
Kinesis
AWS Lambda

Job description

We are seeking a highly skilled and experienced Data Engineering Lead with strong expertise in AWS data services and retail data ecosystems. In this role, you will lead the design, development, and optimization of scalable data pipelines responsible for ingesting and transforming data from MMS (Merchandise Management Systems) and POS (Point of Sale) systems into a centralized Operational Data Store (ODS) to support downstream applications and analytics use cases.

You will play a critical role in building a robust, high-performance data platform using AWS-native services, ensuring data quality, reliability, and real-time or near real-time availability for business operations.

Responsibilities
  • Design and implement scalable ETL/ELT pipelines to ingest data from MMS / POS or third practice systems into AWS-based data platforms.
  • Build and maintain a centralized Operational Data Store (ODS) using Couchbase or similar NoSQL technologies, optimized for low-latency application access.
  • Develop and optimize data processing workflows using Apache Spark / PySpark and AWS services such as AWS Glue and Amazon EMR.
  • Create denormalized, API-ready data models aligned with downstream microservices and application consumption patterns.
  • Implement idempotent processing, CDC merge (upsert) strategies, and data reconciliation mechanisms to ensure consistency across batch and streaming pipelines.
  • Leverage Amazon S3, Amazon Redshift, or Amazon Aurora for efficient storage and querying of structured and semi-structured data.
  • Implement data ingestion patterns (batch and streaming) using tools like Amazon Kinesis or AWS Lambda where applicable.
  • Apply performance tuning and optimization techniques to improve pipeline efficiency, scalability, and cost-effectiveness.
  • Define and enforce data governance, data quality, and metadata management standards across the data platform.
  • Collaborate with DevOps teams to design and maintain CI/CD pipelines using tools like AWS CodePipeline, GitLab, or similar.
  • Conduct peer reviews and provide technical leadership and mentorship to the data engineering team.
  • Collaborate with microservices and application teams to ensure seamless integration with the ODS via APIs and event streams.
Requirements
  • 5+ years of experience in data engineering, with at least 2+ years in a lead role.
  • Strong hands-on experience with AWS data services such as AWS Glue, Amazon S3, Amazon Redshift, and Amazon EMR.
  • Proficiency in PySpark / Spark for large-scale data processing and optimization.
  • Experience designing and implementing ODS layers using NoSQL databases, preferably Couchbase, or similar (DynamoDB, MongoDB).
  • Strong expertise in ETL/ELT design patterns, data ingestion, and transformation pipelines.
  • Experience working with MMS, POS, or retail transaction data is highly preferred.
  • Hands-on experience with CI/CD pipelines (GitLab, AWS CodePipeline, or similar).
  • Good understanding of data governance, data quality, and metadata frameworks.
  • Experience supporting microservices architectures with data platforms.
  • Strong problem-solving skills and ability to optimize complex data workflows.
  • Excellent communication and stakeholder management skills.

Ability to work in a fast-paced, agile environment and manage multiple priorities.


Thanks

Richa

9811509082

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Data Engineer - ETL
Lead Data Engineer - ETL

Srijan Technologies PVT LTD • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Sr. Data Engineer
Sr. Data Engineer

Minfy Technologies • Gurugram District

On-site
INR 1,800,000 - 2,400,000
Sr. Data Engineer (Consultant)
Sr. Data Engineer (Consultant)

Minfy • Chennai District

On-site
INR 1,800,000 - 2,800,000
Lead AWS Data Engineer
Lead AWS Data Engineer

Weekday (YC W21) • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Lead Data Engineer - AWS Glue
Lead Data Engineer - AWS Glue

NetConnectGlobal • Delhi

On-site
INR 1,500,000 - 2,000,000
Data Engineer
Data Engineer

Minfy • India

On-site
INR 1,500,000 - 2,300,000
Data Engineer (AWS)
Data Engineer (AWS)

Weekday (YC W21) • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Data Engineer
Data Engineer

Dentsu Global Services • Mumbai

On-site
INR 900,000 - 1,500,000
Data Engineer, Amazon Last Mile - Routing and Planning - DE
Data Engineer, Amazon Last Mile - Routing and Planning - DE

Amazon • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Acldigital • Pune District

On-site
INR 1,000,000 - 1,500,000