Data Engineer (PySpark, Apache Spark, Delta Lake, ETL, Oracle, DB2, AWS, TWS, Lambda, IAM)

NOVACLOUD SYSTEMS PTE. LTD.

Singapore

On-site

SGD 110,000 - 170,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

NOVACLOUD SYSTEMS PTE. LTD. is seeking a data engineer to design, develop and maintain scalable data pipelines for structured and semi-structured data. You will leverage Databricks, PySpark, and Apache Spark for large-volume processing and build robust ETL/ELT workflows across AWS.

Expertise in SQL and Python with a focus on performance, reliability, and data quality is required, along with experience in Delta Lake and Delta Live Tables where available.

Qualifications

  • 5+ years of experience in data engineering, ETL development, data integration, or related field.
  • Strong hands-on Databricks and data engineering workloads.
  • Proficient in PySpark and Apache Spark for distributed processing.
  • Experience building ETL/ELT pipelines and data solutions.
  • Advanced SQL skills with complex queries, joins, and performance tuning.
  • Hands-on experience with AWS data engineering services.
  • Familiar with Delta Lake and Delta Live Tables preferred.

Responsibilities

  • Design, develop, and maintain scalable data pipelines for structured and semi-structured data.
  • Build ETL/ELT workflows using modern platforms and enterprise tools.
  • Transform data with SQL and Python emphasizing performance and accuracy.
  • Work with Delta Lake and Delta Live Tables for reliable pipelines.
  • Optimize Spark jobs and queries to boost processing efficiency.
  • Collaborate with cross-functional teams to meet data requirements.

Skills

Databricks
PySpark
Apache Spark
SQL
Python
AWS
Delta Lake
Delta Live Tables
IBM DataStage
Informatica PowerCenter
Oracle DB
DB2
Lambda
Redshift
IAM
CloudWatch
Glue
EC2
CI/CD
Git
Jenkins
Agile SDLC

Tools

Databricks
PySpark
Spark
SQL
Python
AWS Cloud
Delta Lake
DLT
DataStage
Informatica PowerCenter
Oracle
DB2
Lambda
Redshift
IAM
Glue
CloudWatch
EC2
GIT
Jenkins

Job description

Responsibilities
  • Design, develop, and maintain scalable data pipelines for ingesting, transforming, validating, and delivering structured and semi-structured data.
  • Develop data engineering solutions using Databricks, PySpark, and Apache Spark for large-volume data processing and transformation.
  • Build and manage ETL/ELT workflows using modern data engineering platforms as well as enterprise ETL technologies.
  • Develop data transformation and processing logic using SQL and Python, with a focus on performance, reliability, and data accuracy.
  • Work with Delta Lake and Delta Live Tables to develop and maintain reliable data pipelines and curated data layers.
  • Develop and optimize ETL workflows using IBM DataStage and Informatica PowerCenter where required.
  • Work with relational databases including Oracle and IBM DB2 for data extraction, transformation, loading, querying, and performance optimization.
  • Implement data integration solutions across cloud and enterprise data platforms.
  • Develop and support data pipelines and associated services on AWS, including data storage, processing, monitoring, and related cloud services.
  • Perform data validation, reconciliation, quality checks, and troubleshooting to ensure the accuracy and completeness of datasets.
  • Optimize Spark jobs, SQL queries, ETL workflows, and data pipelines to improve processing efficiency and overall performance.
  • Collaborate with technical and business teams to understand data requirements and translate them into scalable data solutions.
  • Follow established development, deployment, documentation, data governance, and SDLC practices.
  • Monitor production pipelines, investigate failures, and support timely resolution of data processing issues.
Requirements
  • 5+ years of experience in data engineering, ETL development, data integration, or a related field.
  • Strong hands-on experience with Databricks and data engineering workloads.
  • Hands-on experience in PySpark and Apache Spark for distributed data processing.
  • Solid experience developing ETL/ELT pipelines and data engineering solutions.
  • Strong experience SQL skills, including complex queries, joins, aggregations, optimization, and data transformation.
  • Hands on experience with AWS cloud services used for data engineering and data processing.
  • Experience with Delta Lake, Delta Live Tables (DLT) is preferred.
  • Hands-on experience with IBM DataStage and Informatica PowerCenter.
  • Experience working with relational databases such as Oracle and IBM DB2.
  • Hand on experience in Python for data processing, automation, or pipeline development.
  • Experience with data modelling, data warehousing, data quality, and data integration concepts.
  • Hands on experience in Lambda, Redshift, IAM, CloudWatch, Glue, EC2.
  • Experience with production scheduling, pipeline monitoring, troubleshooting, and deployment processes.
  • Experience in version control and CI/CD practices such as TWS, GIT, Jenkins ect.
  • Ability to work effectively in an Agile/SDLC environment and collaborate with cross-functional teams.
  • Data brick certification would be preferred.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer (PySpark, Apache Spark, Delta Lake, ETL, Oracle, DB2, AWS, TWS, Lambda, IAM)
Data Engineer (PySpark, Apache Spark, Delta Lake, ETL, Oracle, DB2, AWS, TWS, Lambda, IAM)

EXASOFT CONSULTING PTE. LTD. • Singapore

On-site
SGD 90,000 - 170,000
Data Engineer (Databricks)
Data Engineer (Databricks)

BASIL TECHNOLOGIES PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Data Engineer (Databricks)
Data Engineer (Databricks)

Percept Solutions Pte Ltd • Singapore

On-site
SGD 90,000 - 120,000
Data Engineer (Databricks)
Data Engineer (Databricks)

PERCEPT SOLUTIONS PTE. LTD. • Singapore

On-site
SGD 80,000 - 110,000
Data Engineer
Data Engineer

Total eBiz Solutions • Singapore

On-site
SGD 180,000 - 240,000
Senior Data Engineer
Senior Data Engineer

KRIS INFOTECH PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Databricks Engineer
Databricks Engineer

K2 PARTNERING SOLUTIONS PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Databricks Data Engineer
Databricks Data Engineer

Unison Consulting Pte. Ltd. • Singapore

On-site
SGD 120,000 - 180,000
Data Engineer
Data Engineer

Merquri • Singapore

On-site
SGD 70,000 - 110,000
Senior Data Engineer
Senior Data Engineer

CADENZA SOLUTIONS PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000