Lead Data Engineer

Minfy

Gurugram District

On-site

INR 2,600,000 - 4,000,000

Full time

43 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Minfy is seeking a senior Data Engineer to design and implement robust data platforms on AWS. You will build scalable batch and streaming pipelines using Glue, EMR, Kinesis, Lambda, and orchestration with MWAA and Step Functions.

You will optimize Redshift and S3 data lakes, implement CDC with DMS, and ensure data cataloguing, security, and governance are embedded in every workflow.

Qualifications

  • Bachelor's degree in CS/IT/Data Analytics.
  • 6-10 years IT experience with 5-8 years building data apps on AWS.
  • Strong SQL and performance tuning for large datasets.
  • Deep expertise in AWS data services and architectures.

Responsibilities

  • Design, build, and maintain scalable batch and streaming data pipelines using AWS Glue, EMR, Kinesis, and Lambda.
  • Model and optimise analytical data stores in Redshift and S3 data lakes with proper partitioning and tuning.
  • Implement CDC and ingestion from operational databases using DMS, Glue connectors, and Kinesis.
  • Orchestrate and monitor pipelines with MWAA, Step Functions, and EventBridge.
  • Own data cataloguing, lineage, and access control with Glue Data Catalog and Lake Formation.
  • Develop and lifecycle-manage data applications on AWS across EC2, EKS, Lambda, Fargate, RDS/Aurora, DynamoDB, S3.
  • Embed data quality, observability, cost optimization, and security controls into pipelines.
  • Collaborate with analysts, data scientists, and stakeholders to translate requirements into production data products.

Skills

AWS data engineering
SQL
Python / PySpark
Data pipelines
ETL & data integration
AWS Glue/EMR/Redshift
Airflow / MWAA
Data lakes / lake formation
Data quality / observability

Education

Bachelor's degree in CS/IT/Data Analytics

Tools

Amazon S3
Amazon Redshift
Amazon Athena
AWS Glue
Amazon EMR
MWAA

Job description

Job Description:

About The Role

We are seeking a highly skilled and experienced Data Engineer to join our dynamic data and analytics team. The ideal candidate will have strong technical skills in AWS data services and will be responsible for designing, developing, and implementing robust, insightful, data-intensive solutions on AWS. This role requires a deep understanding of data engineering, strong SQL skills, and extensive hands‑on experience with services such as Amazon Redshift, Amazon Athena, AWS Glue, Amazon EMR, Amazon Kinesis, AWS DMS, Amazon S3, and AWS orchestration services for data pipelines. You will play a crucial role in building an AWS-native cloud data platform.

Responsibilities
  • Design, build, and maintain scalable batch and streaming data pipelines using AWS Glue, Amazon EMR, Amazon Kinesis, and AWS Lambda.
  • Model and optimise analytical data stores in Amazon Redshift and S3-based data lakes, including partitioning, file-format, and query-performance tuning.
  • Implement CDC and ingestion patterns from operational databases and SaaS sources using AWS DMS, Glue connectors, and Kinesis.
  • Orchestrate and monitor pipeline workflows using Amazon MWAA (Managed Workflows for Apache Airflow), AWS Step Functions, and Amazon EventBridge.
  • Own data cataloguing, lineage, and fine‑grained access control via AWS Glue Data Catalog and AWS Lake Formation.
  • Contribute to the development, deployment, and lifecycle management of data applications on AWS, leveraging services including EC2, Amazon EKS, AWS Lambda, AWS Fargate, Amazon RDS/Aurora, DynamoDB, and Amazon S3.
  • Embed data quality, observability, cost optimisation, and security controls (encryption, IAM least privilege, PII handling) into every pipeline.
  • Collaborate with analysts, data scientists, and business stakeholders to translate requirements into production‑grade data products.
Required Skills And Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Data Analytics, or a related field.
  • 6-10 years of overall IT experience, with 5-8 years of hands‑on experience designing and developing data applications on AWS.
  • Deep expertise in AWS services and architectures, including but not limited to:
  • Compute: EC2, Amazon EKS, Amazon ECS, AWS Lambda, AWS Fargate, AWS Batch.
  • Storage & Databases: Amazon S3, Amazon RDS, Amazon Aurora, Amazon Redshift, DynamoDB, Amazon Keyspaces, Amazon ElastiCache.
  • Data & Analytics: AWS Glue (ETL, Data Catalog, DataBrew), Amazon EMR, Amazon Athena, Amazon Kinesis (Data Streams, Firehose, Managed Service for Apache Flink), Amazon MSK, AWS DMS, AWS Lake Formation, Amazon MWAA, AWS Step Functions, Amazon QuickSight.
  • Operations & Monitoring: Amazon CloudWatch, AWS CloudTrail, AWS X‑Ray, AWS Cost Explorer.
  • Strong SQL skills, including performance tuning on large datasets.
  • Proven ability to translate business requirements into technical solutions.
  • Excellent analytical, problem‑solving, and critical‑thinking skills.
  • Strong communication and interpersonal skills, with the ability to collaborate effectively with technical and non-technical stakeholders.
  • Experience working in an Agile development methodology.
  • Ability to work independently, manage multiple priorities, and meet tight deadlines.
Preferred Skills (Nice To Have)
  • AWS Certified Data Engineer – Associate, or AWS Certified Solutions Architect certification.
  • Proficiency in Python or PySpark for data manipulation and automation.
  • Experience with infrastructure as code (Terraform, AWS CDK, or CloudFormation) and CI/CD for data pipelines.
  • Experience with another hyperscaler (GCP or Azure), demonstrating breadth in data engineering.
  • Exposure to modern data lakehouse table formats such as Apache Iceberg, Delta Lake, or Apache Hudi.
Requirements:
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Data Engineer
Lead Data Engineer

Acldigital • Pune District

On-site
INR 1,500,000 - 2,000,000
Data Engineer
Data Engineer

Minfy • Gurugram District

On-site
INR 900,000 - 1,500,000
Sr Data Engineer Consultant
Sr Data Engineer Consultant

Minfy • Chennai District

On-site
INR 1,500,000 - 2,200,000
Data Engineer
Data Engineer

Minfy • India

On-site
INR 1,500,000 - 2,300,000
Senior AWS Data Engineer
Senior AWS Data Engineer

EXL • Pune District, Gurugram District, Bengaluru

Hybrid
INR 2,500,000 - 6,000,000
AWS Data Engineer
AWS Data Engineer

Objectways • Bengaluru

On-site
INR 1,500,000 - 2,200,000
AWS Data Engineer
AWS Data Engineer

Zorba Consulting • Pune District

On-site
INR 1,200,000 - 2,000,000
Data Engineer
Data Engineer

Acldigital • Pune District

On-site
INR 1,000,000 - 1,500,000
Data Engineer (AWS)
Data Engineer (AWS)

Weekday (YC W21) • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Data Engineer- AWS
Data Engineer- AWS

NECSWS • Mumbai

On-site
INR 1,200,000 - 1,800,000