Senior Data Engineer – AWS

Jobtailor

Gurugram District

On-site

INR 2,400,000 - 6,000,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Jobtailor is seeking a seasoned Data Engineer to design, build, and optimize scalable data pipelines on AWS, leveraging Glue, S3, Athena, MWAA, and Step Functions. You will contribute to a robust data platform by implementing Medallion Architecture and Iceberg tables, ensuring data quality, lineage, and governance.

With 5-8 years of data engineering experience, strong PySpark and SQL skills, and IaC with Terraform, you will collaborate with architects, analysts, and business teams to modernize

Qualifications

  • 5-8 years of experience in Data Engineering with strong AWS exposure.
  • "Hands-on" experience with AWS Glue, Amazon S3, Athena, MWAA (Apache Airflow), Step Functions, and CloudWatch.
  • Strong programming skills in Python and PySpark.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using AWS Glue (PySpark), Amazon S3, AWS Step Functions, and Athena.
  • Build robust ETL/ELT solutions for batch and near real-time data processing.
  • Develop reusable PySpark frameworks and data transformation components.

Skills

PySpark Programming
ETL/ELT Design
SQL Development
Data Warehousing Concepts
Collaboration

Education

Bachelor's degree in CS/Engineering/IT

Tools

AWS Glue
Amazon S3
Athena
MWAA (Airflow)
Step Functions
CloudWatch
Terraform

Job description

Design, develop, and maintain scalable data pipelines using AWS Glue (PySpark), Amazon S3, AWS Step Functions, and Athena
Build robust ETL/ELT solutions for batch and near real-time data processing
Develop reusable PySpark frameworks and data transformation components
Work with architects and business stakeholders to implement data platform requirements
Participate in migration and modernization initiatives from on-premises data platforms to AWS cloud environments
Implement and support enterprise data lakes using Medallion Architecture
Develop and maintain Apache Iceberg tables for storage, schema evolution, and incremental processing
Ensure data quality, lineage, reconciliation, and auditability across the data platform
Contribute to data modeling and optimization for analytical workloads
Develop and maintain Apache Airflow (MWAA) workflows and DAGs
Automate data movement, validation, monitoring, and notification processes
Implement retry mechanisms, dependency management, and failure handling within workflows
Integrate Airflow with AWS Glue, S3, Athena, and downstream applications
Implement AWS security best practices, including IAM roles, KMS encryption, and secrets management
Support infrastructure provisioning and deployment using Terraform
Collaborate with DevOps, infrastructure, and security teams to maintain secure and reliable cloud environments
Assist in configuring VPC endpoints, networking connectivity, and service integrations
Monitor and troubleshoot production data pipelines and workflows
Perform performance tuning of AWS Glue jobs and Spark workloads
Implement logging, monitoring, and alerting using Amazon CloudWatch
Participate in incident resolution, root cause analysis, and continuous improvement initiatives
Ensure adherence to enterprise operational and governance standards
Collaborate with data architects, analysts, application teams, and business users
Participate in code reviews, design discussions, and technical documentation
Mentor junior engineers and share AWS data engineering best practices
Stay current with emerging AWS services and modern data engineering trends

Requirements
  • Bachelor's degree in Computer Science, Engineering, Information Technology, or related discipline
  • 5-8 years of experience in Data Engineering with strong AWS exposure
  • Hands-on experience with AWS Glue, Amazon S3, Athena, MWAA (Apache Airflow), Step Functions, and CloudWatch
  • Strong programming skills in Python and PySpark
  • Experience building large-scale ETL/ELT data pipelines
  • Good understanding of data modelling, data warehousing, and distributed processing concepts
  • Experience with Apache Iceberg and cloud-native data lake architectures
  • Experience using Infrastructure as Code tools such as Terraform
  • Strong SQL development and query optimization skills
  • Experience supporting production-grade data platforms and troubleshooting complex issues
  • Good understanding of AI technologies, agentic systems, Microsoft Copilot, and Claude
  • Experience in Wealth Management, Asset Management, Capital Markets, Investment Banking, or Financial Services highly preferred
  • Experience with portfolio and investment management, securities and holdings data, trade processing, risk and compliance reporting, regulatory reporting, market and reference data, advisor compensation and payout platforms, or client reporting solutions
  • Experience with Oracle databases or ODI migrations is nice to have
  • Knowledge of CI/CD pipelines and DevOps practices is nice to have
  • Experience with data quality and metadata management tools is nice to have
  • Familiarity with enterprise governance and regulatory compliance requirements is nice to have
  • AWS Certification is nice to have
Core Competencies

Demonstrates expertise in designing and maintaining scalable data pipelines using AWS technologies, including AWS Glue and S3, while ensuring data quality and compliance with enterprise standards. Proficient in developing ETL/ELT solutions and collaborating with cross-functional teams to optimize data processing and architecture.

Highest-signal resume keywords
  • AWS Glue
  • PySpark Programming
  • ETL/ELT Solutions
  • Data Modeling
  • Terraform
ATS Optimization Keywords
Hard Skills
  • Data Engineering
  • AWS Step Functions
  • Apache Airflow
  • SQL Development
  • Apache Iceberg
  • Data Warehousing
  • CloudWatch Monitoring
  • Data Quality Management
  • Infrastructure as Code
  • Performance Tuning
Soft Skills
  • Collaboration
  • Mentoring
  • Problem Solving
  • Communication
  • Continuous Improvement
Certifications & Qualifications
  • AWS Certification
Industry Keywords
  • Wealth Management
  • Asset Management
  • Capital Markets
  • Investment Banking
  • Financial Services
  • Regulatory Compliance
  • Risk Reporting
  • Trade ProcessingClient Reporting
  • Securities Data
Tools & Technologies
  • Amazon S3
  • AWS Glue
  • Athena
  • MWAA (Apache Airflow)
  • Terraform
  • CloudWatch
  • VPC
  • CI/CD Pipelines
  • Data Lakes
  • Medallion Architecture
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer – AWS Data Pipeline
Senior Software Engineer – AWS Data Pipeline

Jobtailor • Bengaluru

On-site
INR 1,500,000 - 2,100,000
AWS Data Engineer
AWS Data Engineer

Qtsolv • Hyderabad

On-site
INR 1,200,000 - 1,800,000
AWS Data Engineer (AWS Glue & PySpark)
AWS Data Engineer (AWS Glue & PySpark)

Tata Consultancy Services • Bengaluru

On-site
INR 1,800,000 - 2,600,000
Lead AWS Data Engineer
Lead AWS Data Engineer

Weekday (YC W21) • Bengaluru

On-site
INR 1,000,000 - 1,500,000
AWS Data Lead Engineer
AWS Data Lead Engineer

Sonata Software • Hyderabad, Chennai District, Bengaluru

Hybrid
INR 4,000,000 - 9,000,000
AWS certifications encouraged
GenAI data foundations exposure
Sr. Data Engineer
Sr. Data Engineer

Minfy Technologies • Gurugram District

On-site
INR 1,800,000 - 2,400,000
Sr. Data Engineer (Consultant)
Sr. Data Engineer (Consultant)

Minfy • Chennai District

On-site
INR 1,800,000 - 2,800,000
AWS Data Engineer
AWS Data Engineer

Careernet • Kolkata District, Bengaluru

On-site
INR 900,000 - 1,500,000
Data Engineer(AWS EMR, ECS, Step functions, Dynamo DB, Glue, PySpark
Data Engineer(AWS EMR, ECS, Step functions, Dynamo DB, Glue, PySpark

Experion • Ernakulam, Thiruvananthapuram

Hybrid
INR 2,000,000 - 4,200,000
AWS Data Engineer
AWS Data Engineer

Objectways • Bengaluru

On-site
INR 1,500,000 - 2,200,000