Data Scientist - Big Data & Spark Analytics

Capgemini Singapore Pte Ltd

Singapore

On-site

SGD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Capgemini Singapore Pte Ltd is seeking a Data Scientist – Big Data to design and implement scalable data solutions using Hadoop ecosystem tools within a collaborative team. You will work with Hive, HDFS, Spark, and Python to deliver robust data pipelines and analytics capabilities.

Join Capgemini to tackle complex data challenges, contribute to DevOps practices, and help clients unlock value from data while adopting modern data architectures.

Qualifications

  • 3–10 years of experience in the Big Data ecosystem (Hive, HDFS, YARN, Spark, Spark SQL, PySpark, Scala, Python).
  • Proficiency in Shell scripting and Python.
  • CI/CD Tools: Jenkins, Azure DevOps (ADO), Bitbucket.
  • Monitoring Tools: Grafana or equivalent.
  • File formats: Parquet, ORC, Sequence (HDFS formats).
  • Strong understanding of concurrent software systems, scalability and robustness.
  • Experience in automation and building end-to-end scalable applications.

Responsibilities

  • Design and implement robust Big Data solutions using the Hadoop ecosystem.
  • Develop data processing pipelines and optimize data workflows.
  • Create automation for builds, testing, and configurations; monitor resource usage.
  • Contribute to data warehousing and data modeling efforts.
  • Implement CI/CD with Bitbucket/Jenkins/ADO for automated deployments.
  • Work with integration tools like FileIT or MQ to enable data flows.
  • Develop scheduling workflows using Control-M and manage resources.

Skills

Spark (PySpark)
Python
Scala
Shell scripting
CI/CD
Hadoop/HDFS
Data modeling

Tools

Jenkins
Azure DevOps
Bitbucket
Control-M
Dataiku

Job description

Capgemini Singapore Pte Ltd is seeking a Data Scientist – Big Data to design and implement scalable data solutions using Hadoop ecosystem tools within a collaborative team. You will work with Hive, HDFS, Spark, and Python to deliver robust data pipelines and analytics capabilities.

Join Capgemini to tackle complex data challenges, contribute to DevOps practices, and help clients unlock value from data while adopting modern data architectures.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Big Data Engineer: Build Scalable Pipelines
Big Data Engineer: Build Scalable Pipelines

Capgemini Singapore Pte Ltd • Singapore

On-site
SGD 60,000 - 90,000
Data Scientist – Big Data
Data Scientist – Big Data

Capgemini Singapore Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
Senior Java & Big Data Architect - Spark, Kafka & Hadoop
Senior Java & Big Data Architect - Spark, Kafka & Hadoop

Capgemini Singapore Pte Ltd • Singapore

On-site
SGD 110,000 - 170,000
Big Data Developer
Big Data Developer

Capgemini Singapore Pte Ltd • Singapore

On-site
SGD 60,000 - 90,000
Hadoop/ BIG Data Developer
Hadoop/ BIG Data Developer

Trinity Workforce Solutions, Inc. • Singapore

On-site
SGD 110,000 - 170,000
Senior Hadoop & Spark Big Data Engineer
Senior Hadoop & Spark Big Data Engineer

Trinity Workforce Solutions, Inc. • Singapore

On-site
SGD 110,000 - 170,000
Senior Big Data Developer: Spark, Hadoop & Microservices
Senior Big Data Developer: Spark, Hadoop & Microservices

Peoplebank Singapore Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
Senior Java Developer with Big Data Experience
Senior Java Developer with Big Data Experience

Capgemini Singapore Pte Ltd • Singapore

On-site
SGD 110,000 - 170,000
Data & Analytics Lead: Architect & Scale Data Solutions
Data & Analytics Lead: Architect & Scale Data Solutions

Capgemini Singapore Pte Ltd • Singapore

On-site
SGD 100,000 - 140,000
Senior Hadoop & Spark Big Data Engineer
Senior Hadoop & Spark Big Data Engineer

Tech Aalto Pte. Ltd • Singapore

On-site
SGD 90,000 - 140,000