Senior Big Data Engineer (Scala/Python + Spark)

Grid Dynamics

Fatih

On-site

TRY 1,072,041 - 1,500,857

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Opportunity to work on bleeding-edge projects
Competitive salary
Flexible schedule
Professional development opportunities

Job summary

A global technology company is seeking a Mid-Senior Level Engineer to develop a high-performance data analytics platform. Key responsibilities include creating scalable data ingestion jobs using Scala and Apache Spark, implementing robust ETL pipelines, and maintaining high data quality. Candidates must have strong experience with cloud storage solutions and distributed systems. The role offers competitive pay, professional growth opportunities, and the chance to work on cutting-edge projects in a collaborative environment.

Qualifications

  • Expert-level proficiency with Apache Spark for batch and streaming.
  • Hands-on experience with Apache Iceberg or similar formats.
  • Strong engineering skills in Scala and/or Python.

Responsibilities

  • Develop and maintain scalable data ingestion and transformation jobs.
  • Implement Spark-based ETL/ELT pipelines for data migration.
  • Ensure data quality checks and validations throughout the lifecycle.

Skills

Apache Spark
Scala
Python
Apache Iceberg
Kubernetes
Terraform
HDFS

Tools

Apache Flink
AWS S3

Job description

Our customer is one of the world’s largest technology companies based in Silicon Valley with operations all over the world. In this project, we are working on the bleeding edge of Big Data technology to develop a high-performance data analytics platform, which handles petabytes of data.

Responsibilities
  • Develop, implement, and maintain scalable data ingestion and transformation jobs using Scala/Python.
  • Implement robust Spark-based ETL/ELT pipelines to migrate data efficiently from legacy systems (HDFS/Hive) to modern cloud storage solutions.
  • Apply rigorous data quality checks and validation processes throughout the migration lifecycle.
  • Participate actively in code reviews, ensuring adherence to the team's best practices and writing clean, testable, and maintainable code.
  • Document technical designs, pipeline logic, and standard operational procedures.
  • Support troubleshooting, debugging, and bug fixing during critical migration and deployment activities.
  • Contribute to AI engineering or Prompt engineering efforts related to data platform usage.
Requirements
  • Expert-level proficiency with Apache Spark (batch and streaming).
  • Hands‑on experience with Apache Iceberg (or similar formats like Delta Lake or Apache Hudi) for implementing ACID transactions, schema evolution, and time‑travel capabilities.
  • Deep knowledge of HDFS internals and large‑scale migration strategies.
  • Strong engineering skills in Scala and/or Python.
  • Experience running Spark and/or Flink jobs on Kubernetes (e.g., using the Spark‑on‑K8s operator).
  • Experience with distributed blob storages (e.g., AWS S3, Ceph, etc.).
  • Proven ability to build ingestion, transformation, and enrichment pipelines for complex, large‑scale datasets.
  • Familiarity with Infrastructure‑as‑Code tools like Terraform or Helm for provisioning and managing data infrastructure.
Nice to have
  • Experience with Apache Flink for high‑velocity streaming data processing.
  • Prior hands‑on experience in major migration projects or large‑scale data platform modernization initiatives.
We offer
  • Opportunity to work on bleeding‑edge projects
  • Work with a highly motivated and dedicated team
  • Competitive salary
  • Flexible schedule
  • Professional development opportunities
About Us

Grid Dynamics (NASDAQ: GDYN) is a leading provider of technology consulting, platform and product engineering, AI, and advanced analytics services. Fusing technical vision with business acumen, we solve the most pressing technical challenges and enable positive business outcomes for enterprise companies undergoing business transformation. A key differentiator for Grid Dynamics is our 8 years of experience and leadership in enterprise AI, supported by profound expertise and ongoing investment in data, analytics, cloud & DevOps, application modernization and customer experience. Founded in 2006, Grid Dynamics is headquartered in Silicon Valley with offices across the Americas, Europe, and India.

Seniority level: Mid‑Senior level

Employment type: Full‑time

Job function: Engineering and Information Technology

Industries: IT Services and IT Consulting

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Big Data Engineer (Scala + Spark)
Senior Big Data Engineer (Scala + Spark)

Grid Dynamics • Fatih

On-site
TRY 1,269,000 - 2,116,000
Opportunity to work on bleeding-edge projects
Competitive salary
Flexible schedule
+1
Senior Big Data Engineer - Spark/Scala-Python, Flexible
Senior Big Data Engineer - Spark/Scala-Python, Flexible

Grid Dynamics • Fatih

On-site
TRY 1,072,000 - 1,501,000
Senior Data Engineer - Spark, Scala & Dashboards (Remote)
Senior Data Engineer - Spark, Scala & Dashboards (Remote)

Grid Dynamics • Fatih

On-site
TRY 1,269,000 - 2,116,000
Senior AI Engineer with Databricks
Senior AI Engineer with Databricks

EPAM Systems • Turkey

On-site
TRY 250,000 - 450,000
Private health insurance
Continuous upskilling
English courses
+2
Senior Data Engineer
Senior Data Engineer

OREDATA • Turkey

On-site
TRY 400,000 - 640,000
Senior Big Data Administrator
Senior Big Data Administrator

Ithinka IT and IoT Technologies • Turkey

On-site
TRY 450,000 - 750,000
Lead Data Software Engineer with Databricks, Apache Kafka, Apache Spark, Kubernetes
Lead Data Software Engineer with Databricks, Apache Kafka, Apache Spark, Kubernetes

EPAM Systems, Inc. • Turkey

On-site
TRY 3,848,000 - 5,772,000
Extra leave days
Referral bonuses
Private health insurance
+5
Data Engineer (f/m/x)
Data Engineer (f/m/x)

Cepres GmbH • Fatih

On-site
TRY 60,000 - 80,000
Senior Data Engineer: Real-Time Pipelines & Data Governance
Senior Data Engineer: Real-Time Pipelines & Data Governance

Coderspace • Fatih

On-site
TRY 350,000 - 550,000
Private health insurance
Meal card
Monthly IstanbulCard reimbursement
+1
Senior Data Engineer
Senior Data Engineer

Personify Health • Tuzla

On-site
TRY 400,000 - 700,000