Data & AI Solutions Specialist

REGTECH INSIGHT PTE. LTD.

Singapore

On-site

SGD 150,000 - 210,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

REGTECH INSIGHT PTE. LTD. in Singapore is seeking a Senior Data Engineer to design, develop and optimize scalable data pipelines using Spark, Hadoop and cloud technologies. You will collaborate with architects, data scientists and engineers to deliver production-ready solutions.

The role emphasizes real-time and batch processing, data warehouses and data lakes, and integrating AI capabilities. Strong Python/Java/Scala development and AWS experience are essential for success.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Information Technology, Engineering or a related field.
  • Minimum 8 years of relevant experience in Data Engineering / Big Data.
  • Strong programming experience in Python, Java and/or Scala.
  • Hands-on experience with Apache Spark/PySpark, SQL, Hadoop and Hive.
  • Experience with cloud platforms, particularly AWS.
  • Experience with data warehouses, relational databases and data lake technologies.
  • Experience with Kafka, Airflow, Jenkins, Git and CI/CD practices.
  • Experience with Docker, Kubernetes or OpenShift is advantageous.
  • Knowledge of Databricks, Snowflake and Cloudera is advantageous.
  • Exposure to Generative AI, LLMs, RAG, LangChain/LangGraph and vector databases is an advantage.
  • Strong analytical, problem-solving and communication skills.

Responsibilities

  • Design, develop and maintain scalable batch and real-time data pipelines using Apache Spark/PySpark, SQL and Python.
  • Develop and optimize ETL/ELT pipelines for large-scale data processing.
  • Build data ingestion solutions using databases, APIs, files and streaming platforms.
  • Work with AWS services including S3, Glue, EMR, Redshift, Kinesis, Lambda and DynamoDB.
  • Develop and optimize data processing solutions using Hadoop, Hive, Spark, Kafka, Cloudera and Databricks.
  • Perform SQL and Spark performance tuning and optimize large-scale data processing workloads.
  • Design and implement data models, data warehouses and data lake solutions.
  • Develop CI/CD pipelines and support containerized deployments using Docker, Kubernetes and OpenShift.
  • Integrate Generative AI and NLP capabilities into enterprise data applications where required.
  • Develop solutions using LLM frameworks, RAG, vector databases and AI APIs.
  • Collaborate with solution architects, data scientists, software engineers and business stakeholders to deliver production-ready solutions.
  • Participate in system design, development, testing, deployment and production support.

Skills

Python
Java
Scala
PySpark
Apache Spark
SQL
Hadoop
Hive
Kafka
Databricks
Cloudera
AWS
S3
Glue
EMR
Redshift
Kinesis
Lambda
Airflow
Jenkins
Docker
Kubernetes
OpenShift
Snowflake
MongoDB
Oracle
PostgreSQL
Generative AI
LangChain

Education

Bachelor's or Master’s degree in CS/IT/Engineering

Tools

Docker
Kubernetes
OpenShift
Airflow
Jenkins
Git
CI/CD
Snowflake
Databricks
Hive
Hadoop
Kafka

Job description

We are looking for a highly skilled and experienced Senior Data Engineer to design, develop and optimize scalable data engineering solutions using Big Data, cloud and modern AI technologies.

The successful candidate will be responsible for developing high-performance data pipelines, data ingestion frameworks and data processing solutions, while working closely with architects, business stakeholders, data scientists and engineering teams.

Key Responsibilities

Design, develop and maintain scalable batch and real-time data pipelines using Apache Spark/PySpark, SQL and Python.

Develop and optimize ETL/ELT pipelines for large-scale data processing.

Build data ingestion solutions using databases, APIs, files and streaming platforms.

Work with AWS services including S3, Glue, EMR, Redshift, Kinesis, Lambda and DynamoDB.

Develop and optimize data processing solutions using Hadoop, Hive, Spark, Kafka, Cloudera and Databricks.

Perform SQL and Spark performance tuning and optimize large-scale data processing workloads.

Design and implement data models, data warehouses and data lake solutions.

Develop CI/CD pipelines and support containerized deployments using Docker, Kubernetes and OpenShift.

Integrate Generative AI and NLP capabilities into enterprise data applications where required.

Develop solutions using LLM frameworks, RAG, vector databases and AI APIs.

Collaborate with solution architects, data scientists, software engineers and business stakeholders to deliver production-ready solutions.

Participate in system design, development, testing, deployment and production support.

Requirements

Bachelor’s or Master’s degree in Computer Science, Information Technology, Engineering or a related field.

Minimum 8 years of relevant experience in Data Engineering / Big Data, with strong experience in large-scale data processing.

Strong programming experience in Python, Java and/or Scala.

Strong hands-on experience with Apache Spark/PySpark, SQL, Hadoop and Hive.

Experience with cloud platforms, particularly AWS.

Experience with data warehouses, relational databases and data lake technologies.

Experience with Kafka, Airflow, Jenkins, Git and CI/CD practices.

Experience with Docker, Kubernetes or OpenShift is advantageous.

Knowledge of Databricks, Snowflake and Cloudera is advantageous.

Exposure to Generative AI, LLMs, RAG, LangChain/LangGraph and vector databases is an advantage.

Strong analytical, problem-solving and communication skills.

Technical Skills

Python, Java, Scala, PySpark, Apache Spark, SQL, Hadoop, Hive, Kafka, Databricks, Cloudera, AWS, S3, Glue, EMR, Redshift, Kinesis, Lambda, Airflow, Jenkins, Docker, Kubernetes, OpenShift, Snowflake, MongoDB, Oracle, PostgreSQL and Generative AI technologies.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data & AI Solutions Specialist
Data & AI Solutions Specialist

Riskdata Consulting • Singapore

On-site
SGD 120,000 - 160,000
Data & AI Solutions Specialist
Data & AI Solutions Specialist

UARROW PTE. LTD. • Singapore

On-site
SGD 120,000 - 160,000
Data & AI Solutions Specialist
Data & AI Solutions Specialist

RISKDATA CONSULTING PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Data & AI Solutions Specialist
Data & AI Solutions Specialist

JEET ANALYTICS PTE. LTD. • Singapore

On-site
SGD 180,000 - 240,000
Data & AI Solutions Specialist
Data & AI Solutions Specialist

UNISONEDGE CONSULTING PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Data & AI Solutions Specialist
Data & AI Solutions Specialist

re-zoo-me • Singapore

Hybrid
SGD 120,000 - 180,000
Senior AI Data Engineer
Senior AI Data Engineer

jobline resources pte. ltd. • Singapore

On-site
SGD 90,000 - 180,000
Senior Data Engineer – PySpark, Databricks & Data Lakehouse
Senior Data Engineer – PySpark, Databricks & Data Lakehouse

D L Resources Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Data Engineer (Ref 26570)
Senior AI Data Engineer (Ref 26570)

JOBLINE RESOURCES PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Senior Data Engineer
Senior Data Engineer

INTELLECT MINDS PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000