Data Engineer | Noida

DigitalXNode

Dadri

On-site

INR 1,200,000 - 1,800,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

DigitalXNode in Noida, India is seeking a Data Engineer to design, develop, and maintain scalable data pipelines and big data solutions. You will collaborate with data scientists, software engineers, business analysts, and product teams to enable analytics, reporting, and BI.

Responsibilities include building ETL/ELT pipelines with Scala, Spark, Hadoop and SQL, optimizing jobs, ensuring data quality, and supporting cloud-ready data architectures.

Qualifications

  • Hands-on experience with Scala, Spark, Hadoop, and SQL.
  • Experience developing ETL/ELT pipelines and data modeling knowledge.
  • Understanding of distributed computing and large-scale data processing.
  • Experience with cloud platforms is a plus.

Responsibilities

  • Design, develop, test, deploy, and maintain scalable data pipelines using Scala, Spark, Hadoop, and SQL.
  • Build efficient ETL/ELT workflows for structured and unstructured data.
  • Develop reusable, optimized data processing components and automate ingestion.
  • Ensure high-quality, reliable data pipelines and monitor performance.
  • Collaborate with data scientists, analysts, and product teams; participate in Agile sprints.

Skills

Scala
Apache Spark
Hadoop
SQL
Python

Education

Bachelor's degree in Computer Science, Information Technology, Data Science, Software Engineering, or related field
B.Tech / BE in Information Technology, Computer Science, Electronics or equivalent
MCA / M.Tech / M.Sc. in CS / Data Science / IT or equivalent (added advantage)

Tools

Docker
Kubernetes
Airflow

Job description

Data Engineer | Noida

We are seeking a motivated and detail-oriented Data Engineer to design, develop, and maintain scalable data pipelines and big data solutions. The ideal candidate should have hands‑on experience with Scala, Apache Spark, Hadoop, SQL, and modern data engineering practices. You will work closely with data scientists, software engineers, business analysts, and product teams to build reliable data platforms that enable analytics, reporting, and business intelligence.

In this role, you will be responsible for developing high‑performance data pipelines, optimizing large‑scale data processing workflows, ensuring data quality, and supporting enterprise data initiatives. This position offers an excellent opportunity to work with modern big data technologies and cloud‑ready data architectures.

Key Responsibilities
Data Pipeline Development
  • Design, develop, test, deploy, and maintain scalable data pipelines using Scala, Apache Spark, Hadoop, and SQL.
  • Build efficient ETL/ELT workflows for processing structured and unstructured data.
  • Develop reusable and optimized data processing components.
  • Automate data ingestion, transformation, and loading processes.
  • Ensure high-quality, reliable, and maintainable data pipelines.
Big Data Engineering
  • Process large-scale datasets using distributed computing frameworks.
  • Optimize Spark jobs for performance, scalability, and resource utilization.
  • Work with Hadoop ecosystem components for distributed data storage and processing.
  • Support batch and near real‑time data processing workflows.
  • Monitor and improve big data platform performance.
Database & Data Management
  • Design and optimize SQL queries for efficient data retrieval.
  • Work with relational databases such as MySQL and PostgreSQL.
  • Ensure data integrity, consistency, and accuracy across systems.
  • Support database optimization and performance tuning.
  • Manage data storage and lifecycle processes.
Collaboration & Solution Delivery
  • Collaborate with cross‑functional teams to understand business and technical requirements.
  • Translate business needs into scalable data engineering solutions.
  • Work closely with data scientists and analysts to prepare datasets for analytics and machine learning.
  • Participate in Agile development processes and sprint planning.
  • Maintain technical documentation and implementation records.
Performance & Quality
  • Troubleshoot complex issues related to data processing, storage, and retrieval.
  • Ensure data platform scalability, security, reliability, and availability.
  • Perform data validation and quality assurance activities.
  • Optimize data workflows for improved efficiency and reduced processing time.
  • Follow data engineering best practices and coding standards.
Required Skills
Data Engineering
  • Strong experience in Data Engineering concepts and practices.
  • Hands‑on experience with:
    • Scala
    • Apache Spark
    • Hadoop
    • SQL
  • Understanding of distributed computing and large‑scale data processing.
  • Experience developing ETL/ELT pipelines.
  • Knowledge of data modeling and data architecture principles.
Big Data Technologies
  • Hadoop Ecosystem
  • HDFS
  • MapReduce
  • Apache Spark
  • Distributed Data Processing
  • Batch Processing
Databases
  • MySQL
  • PostgreSQL
  • SQL Query Optimization
  • Database Performance Tuning
  • Data Modeling
Programming & Development
  • Scala
  • SQL
  • Python (Preferred)
  • Git & Version Control
  • Linux Basics
Professional Skills
  • Strong analytical and problem‑solving skills.
  • Excellent communication and collaboration abilities.
  • Ability to work independently and in Agile teams.
  • Strong attention to detail and commitment to quality.
  • Willingness to learn emerging data engineering technologies.
Preferred Skills
  • Experience with cloud platforms such as AWS, Azure, or Google Cloud Platform.
  • Familiarity with Apache Kafka or streaming technologies.
  • Knowledge of Airflow or workflow orchestration tools.
  • Experience with Docker and Kubernetes.
  • Exposure to Data Warehousing concepts.
  • Understanding of CI/CD pipelines and DevOps practices.
  • Basic knowledge of machine learning data pipelines.
Technologies & Tools
Big Data
  • Apache Spark
  • Hadoop
  • HDFS
  • MapReduce
Programming
  • Scala
  • SQL
  • Python
Databases
  • MySQL
  • PostgreSQL
Cloud & DevOps
  • AWS (Preferred)
  • Azure (Preferred)
  • Docker
  • Kubernetes
Development Tools
  • Git
  • Linux
  • IntelliJ IDEA
  • Maven
Education
  • Bachelor's degree in Computer Science, Information Technology, Data Science, Software Engineering, or a related field.
  • B.Tech / BE in Information Technology, Computer Science, Electronics, or equivalent disciplines preferred.
  • MCA, M.Tech, M.Sc. (Computer Science, Data Science, Information Technology), or equivalent qualifications are an added advantage.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Advance Career Solutions • Pune District, Chennai District, Bengaluru

On-site
INR 1,200,000 - 2,800,000
Data Engineer
Data Engineer

Crayon Data • Chennai District

On-site
INR 800,000 - 1,200,000
Data Engineer
Data Engineer

e-Stone Information Technology Private Limited • Mumbai

On-site
INR 1,500,000 - 2,800,000
Data Engineer
Data Engineer

Synergy Computer Solutions • Hyderabad

On-site
INR 1,200,000 - 2,000,000
Senior Data Engineer
Senior Data Engineer

Quess IT Solutions • Bengaluru

On-site
INR 3,000,000 - 6,000,000
Data Engineer
Data Engineer

Meril • Vapi

On-site
INR 800,000 - 1,200,000
Data Engineer
Data Engineer

Impronics Technologies • Bengaluru

On-site
INR 700,000 - 1,200,000
Competitive compensation and benefits package
Opportunities for professional growth
Collaborative and innovative work environment
Data Engineer
Data Engineer

Altysys • Gurugram District

On-site
INR 800,000 - 1,200,000
Data Engineer - Consultant
Data Engineer - Consultant

Iris Software • Dadri

On-site
INR 1,400,000 - 2,200,000
Data Engineer
Data Engineer

ConveGenius • Chennai District

On-site
INR 800,000 - 1,200,000