Immediate Databricks Infrastructure

ITCAN PTE. LIMITED

Singapore

On-site

SGD 70,000 - 110,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

ITCAN PTE. LIMITED is seeking a Data Engineer to design, develop, and maintain scalable data pipelines on Databricks and cloud platforms.

You will integrate diverse data sources and support analytics, reporting, and ML workloads while upholding governance and reliability. You will collaborate with analytics, product, and infrastructure teams to advance the data platform, implement monitoring, and optimise performance using Azure, Delta Lake, and related tools.

Qualifications

  • 3+ years of data engineering experience with scalable pipelines.
  • Experience with PySpark, Spark SQL, Databricks notebooks/jobs.
  • Experience with Azure data services and orchestrators like ADF or Airflow.

Responsibilities

  • Design, develop, and maintain scalable data pipelines on Databricks and cloud platforms.
  • Integrate data from databases, APIs, log files, and streaming platforms.
  • Develop data transformation routines to clean, normalize, and aggregate data.
  • Implement data governance in alignment with company standards.
  • Collaborate with analytics, product, and infrastructure teams to enhance data platform.

Skills

Data engineering
ETL pipelines
PySpark
SQL
Databricks
Azure
Airflow
CI/CD
Scrum
Kafka/Flink

Tools

ADF
Airflow
Git

Job description

Job Description

The Data Engineer will be responsible for designing, developing, and maintaining scalable and reliable data pipelines on Databricks and cloud platforms. The role requires integrating diverse data sources, ensuring high-quality data processing, and supporting analytics, reporting, and machine learning workloads. The role involves collaborating closely with analytics, product, and infrastructure teams to enhance the company’s data platform while adhering to best practices for governance, monitoring, and reliability.

What will you do?
  • Develop and maintain ETL pipelines for centralized data storage systems (e.g. Delta Lake).
  • Integrate data from databases, APIs, log files, streaming platforms, and external providers
  • Develop data transformation routines to clean, normalize, and aggregate data
  • Apply data processing techniques to handle complex or inconsistent datasets
  • Contribute to frameworks and best practices for code development and deployment
  • Implement data governance in alignment with company standards
  • Partner with analytics and product leaders to design and operationalize pipelines
  • Collaborate with infrastructure leaders to advance cloud-based data platforms
  • Explore new tools and techniques leveraging Azure, Databricks, or related platforms
  • Monitor data pipelines to detect and resolve issues promptly
  • Develop monitoring tools, alerts, and automated error-handling mechanisms
  • Analyse business requirements and identify data extraction requirements
  • Attend requirement grooming and refinement sessions with users
  • Develop and maintain ETL pipelines for ingestion, transformation, validation, and loading
  • Optimise performance and batch scheduling
  • Develop dashboards, reports, scorecards, and data visualizations
  • Perform SIT, data validation, data profiling and confirm data accuracy
  • Validate completeness and consistency of ETL Loads
  • Support UAT and production implementation
Qualifications

The ideal candidate should possess:

  • 3 or more years of experience in data engineering with scalable pipelines
  • Strong experience designing data solutions including data modelling and distributed computing architectures
  • Hands-on experience with data processing jobs using PySpark, Spark SQL, and Databricks notebooks/jobs
  • Experience orchestrating data pipelines with ADF, Airflow, or similar tools
  • Experience with both real-time and batch data processing
  • Experience building pipelines on Azure, with AWS experience beneficial
  • Proficiency in SQL including window functions and performance optimization
  • Understanding of DevOps tools, Git workflows, and CI/CD pipelines
  • Familiarity with Scrum methodology and experience working in Scrum teams
  • Ability to apply Scrum practices in a practical project context
  • Strong problem-solving and collaborative mindset
  • Experience with streaming technologies such as Apache Kafka, Apache Flink, or AWS Kinesis
  • Ability to design and implement real-time data processing pipelines
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer (Databricks)
Data Engineer (Databricks)

BASIL TECHNOLOGIES PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Databricks Engineer
Databricks Engineer

ESPIRE INFOLABS (SINGAPORE) PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Databricks Engineer
Databricks Engineer

K2 PARTNERING SOLUTIONS PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Data Engineer (Databricks)
Data Engineer (Databricks)

Percept Solutions Pte Ltd. • Singapore

On-site
SGD 90,000 - 120,000
Data Engineer
Data Engineer

Epergne Solutions • Singapore

On-site
SGD 120,000 - 180,000
DATA ENGINEER
DATA ENGINEER

REGTECH INSIGHT PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Information Technology - Lead Data Engineer
Information Technology - Lead Data Engineer

Singapore Airlines • Singapore

On-site
SGD 80,000 - 120,000
Data Engineer
Data Engineer

UNISYNC SYSTEMS PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Data Engineer
Data Engineer

UNISONEDGE CONSULTING PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Data Engineer
Data Engineer

UARROW PTE. LTD. • Singapore

On-site
SGD 110,000 - 190,000