Data Engineer

Aligned Automation

Maharashtra

On-site

INR 800,000 - 1,200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading technology solutions provider in India is seeking a skilled Data Engineer to design and optimize data pipelines using PySpark and Apache Airflow. The role involves managing workflow orchestration, ensuring data quality, and collaborating with data teams. Candidates should have strong SQL skills and experience in cloud platforms like AWS or Azure. This position offers the opportunity to contribute to impactful projects in a vibrant work environment.

Qualifications

  • Hands-on experience with PySpark for data processing.
  • Experience with Apache Airflow for workflow orchestration.
  • Strong SQL skills for managing large datasets.

Responsibilities

  • Design and maintain ETL/ELT pipelines using PySpark.
  • Manage workflows with Apache Airflow.
  • Optimize data pipelines for performance.

Skills

PySpark
Apache Airflow
SQL
Spark architecture
Data warehousing concepts
Git
REST APIs

Education

Bachelor’s or Master’s degree in Computer Science, IT, Engineering, or related field

Tools

AWS
Azure
GCP
Docker
Kubernetes

Job description

At Aligned Automation, we live by our "Better Together" philosophy to build a better world. As a strategic service provider to Fortune 500 companies, we help digitize enterprise operations and drive impactful business strategies. Our purpose goes beyond projects—we strive to deliver meaningful, sustainable change that shapes a more optimistic and equitable future.

Our culture is deeply rooted in our 4Cs—Care, Courage, Curiosity, and Collaboration—ensuring that each employee is empowered to grow, innovate, and thrive in an inclusive workplace.

Job Summary

We are seeking a skilled Data Engineer with strong expertise in PySpark and Apache Airflow to design, build, and optimize scalable data pipelines. The ideal candidate should have experience in big data processing, workflow orchestration, and cloud-based data platforms.

Key Responsibilities
  • Design, develop, and maintain scalable ETL/ELT pipelines using PySpark
  • Build and manage workflow orchestration using Apache Airflow
  • Process large datasets using distributed computing frameworks (Spark)
  • Optimize data pipelines for performance, reliability, and scalability
  • Implement data quality checks and monitoring mechanisms
  • Work closely with Data Analysts, Data Scientists, and BI teams
  • Manage data ingestion from various sources (APIs, databases, flat files, streaming)
  • Troubleshoot and resolve pipeline failures
  • Implement CI/CD for data pipelines
  • Ensure data governance and security best practices
Required Skills
  • Strong hands‑on experience in PySpark
  • Experience in Apache Airflow (DAGs, Operators, Scheduling)
  • Good understanding of Spark architecture
  • Strong SQL knowledge
  • Experience with data warehousing concepts
  • Experience with:
    • S3 / ADLS / GCS
    • Redshift / Snowflake / BigQuery
  • Knowledge of Git and version control
  • Understanding of REST APIs and data ingestion
Good to Have
  • Experience with cloud platforms (AWS / Azure / GCP)
  • Experience with Kafka or streaming pipelines
  • Docker & Kubernetes knowledge
  • Delta Lake / Iceberg knowledge
  • Experience in CI/CD tools (Jenkins, GitHub Actions)
  • Experience in monitoring tools (Prometheus, Grafana)
Educational Qualification
  • Bachelor’s or Master’s degree in Computer Science, IT, Engineering, or related field
Soft Skills
  • Strong problem‑solving skills
  • Good communication and collaboration skills
  • Ability to work in an agile environment
  • Ownership mindset and attention to detail
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer_Spark/Scala
Data Engineer_Spark/Scala

Zorba AI • Kolkata District

On-site
INR 1,000,000 - 1,500,000
Data Engineer_Spark/Scala
Data Engineer_Spark/Scala

Zorba AI • Maharashtra

On-site
INR 800,000 - 1,200,000
Data Engineer_Spark/Scala
Data Engineer_Spark/Scala

Zorba AI • Mumbai

On-site
INR 1,000,000 - 1,500,000
Data Engineer Pyspark and Mongo DB
Data Engineer Pyspark and Mongo DB

Aligned Automation, LLC • Pune District

On-site
INR 1,500,000 - 2,300,000
Data Engineer
Data Engineer

Advance Career Solutions • Pune District, Chennai District, Bengaluru

Hybrid
INR 1,200,000 - 2,800,000
Data Engineer
Data Engineer

Team Computers • Mumbai, Ahmedabad District

On-site
INR 1,800,000 - 2,400,000
Pyspark developer
Pyspark developer

Aligned Automation • Pune District

On-site
INR 2,200,000 - 3,400,000
Assistant Manager - Data Engineer
Assistant Manager - Data Engineer

Kavi India • Chennai District

On-site
INR 1,000,000 - 1,700,000
Sr. Data Engineer (ETL/ELT)
Sr. Data Engineer (ETL/ELT)

NexTurn Inc. • Hyderabad

On-site
INR 1,000,000 - 1,500,000
Senior Data Engineer
Senior Data Engineer

Moolya Software Testing • Bengaluru

On-site
INR 1,500,000 - 2,300,000