Senior Data Engineer

Datazza

Fatih

On-site

TRY 500,000 - 900,000

Full time

25 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Datazza is a data engineering and analytics company building scalable data platforms for enterprise clients. We are seeking a Senior Data Engineer to join a large-scale telecom data platform project, working with big data and lakehouse environments, and handling batch and streaming pipelines.

You will design, develop, and optimize ELT/ETL, data models, data marts, and analytics layers using Apache Spark, Kafka, Iceberg/Delta Lake/Hudi, and distributed query engines.

Qualifications

  • 5+ years of experience in data engineering, big data, or data warehousing.
  • Production-grade data pipelines experience.
  • Advanced SQL, complex transformations, performance tuning.
  • Strong Python/Scala/Java for data processing.
  • Experience with Spark and Kafka in distributed environments.
  • Knowledge of lakehouse architectures (Iceberg/Delta/Hudi).

Responsibilities

  • Design, develop, and maintain scalable batch and real-time data pipelines.
  • Build and optimize data processing workloads for telecom data.
  • Develop ELT/ETL across lake, lakehouse, and enterprise DW.
  • Design analytical data models, data marts, and reusable layers.
  • Work with Spark and Kafka for processing and streaming.
  • Implement lakehouse structures using Iceberg/Delta/Hudi.
  • Develop SQL workloads on Trino, Spark SQL, or similar.
  • Integrate relational, NoSQL, APIs, and object storage.
  • Perform query optimization, partitioning, and tuning.
  • Implement data quality controls, monitoring, and observability.

Skills

Big data
Data pipelines
SQL
Python
Spark
Kafka
Data modeling
ETL/ELT
Distributed systems
Troubleshooting
Team collaboration

Education

Bachelor's or Master's in CS/CE/IS

Tools

Apache Spark
Apache Kafka
Airflow
Iceberg
Delta Lake
Hudi
Trino
Presto
Docker
Kubernetes
Git

Job description

Employment Type: Full-time, On-site

Open Positions: 1

Datazza is a specialized data engineering and advanced analytics company that helps organizations build modern, scalable, and secure data platforms.

Our team consists of experienced data professionals with more than 15 years of hands-on expertise in enterprise data warehouse, big data, analytics, and data platform projects. We design and implement on‑premise and cloud‑based data platforms, lakehouse architectures, real‑time and batch data pipelines, analytical data products, machine learning platforms, and business intelligence solutions.

We primarily work with open‑source and cloud‑native technologies to develop flexible, vendor‑independent, and production‑grade data platforms. As an agile and engineering‑focused company, we work closely with our clients and take an active role in architectural design, implementation, optimization, and operationalization.

Role Overview

We are looking for a Senior Data Engineer to join our team and work on a large‑scale data platform project in the telecommunications industry.

In this role, you will design, develop, and optimize high‑volume data pipelines and analytical data structures operating within modern big data and lakehouse environments. You will work with large‑scale batch and streaming datasets generated by telecommunications systems and contribute to the development of scalable, reliable, and high‑performance data platforms.

The ideal candidate has strong hands‑on data engineering experience, understands distributed data processing concepts, and is comfortable working with open‑source technologies in enterprise production environments.

Key Responsibilities
  • Design, develop, and maintain scalable batch and real‑time data pipelines.
  • Build and optimize data processing workloads for large‑volume telecommunications datasets.
  • Develop ELT and ETL processes across data lake, lakehouse, and enterprise data warehouse environments.
  • Design analytical data models, data marts, and reusable data layers.
  • Work with distributed processing and streaming technologies such as Apache Spark and Apache Kafka.
  • Implement lakehouse data structures using technologies such as Apache Iceberg, Delta Lake, or Apache Hudi.
  • Develop and optimize SQL workloads running on distributed query engines such as Trino, Spark SQL, or similar platforms.
  • Integrate relational databases, NoSQL systems, event streams, object storage, APIs, and enterprise data sources.
  • Perform query optimization, data partitioning, file sizing, compaction, and performance tuning.
  • Implement data quality controls, monitoring, logging, lineage, and observability practices.
  • Contribute to architectural decisions, technology evaluations, and technical standards.
  • Produce technical documentation, data mappings, and operational runbooks.
  • Collaborate with architects, analysts, data scientists, DevOps teams, and client stakeholders.
  • Review code, promote engineering best practices, and mentor less‑experienced team members.
Required Qualifications
  • At least 5 years of professional experience in data engineering, big data, data warehouse, or related roles.
  • Strong hands‑on experience in building and operating production‑grade data pipelines.
  • Advanced SQL skills, including complex transformations, query optimization, and performance tuning.
  • Strong experience with Python, Scala, or Java for data processing and automation.
  • Hands‑on experience with Apache Spark or a comparable distributed data processing framework.
  • Experience with Apache Kafka or another event‑streaming platform.
  • Strong understanding of data warehouse, data lake, and lakehouse architecture principles.
  • Experience with dimensional modeling, normalized data models, data marts, and analytical data structures.
  • Familiarity with workflow orchestration technologies such as Apache Airflow.
  • Experience integrating data from relational databases, APIs, files, object storage, and event‑based systems.
  • Good understanding of distributed systems, data partitioning, parallel processing, and fault tolerance.
  • Experience with Linux‑based environments, Git, CI/CD processes, and software development practices.
  • Strong analytical thinking, troubleshooting, communication, and problem‑solving skills.
  • Ability to work collaboratively with technical teams and client stakeholders.
  • Bachelor's or master's degree in Computer Science, Computer Engineering, Information Systems, or a related field, or equivalent practical experience.
Preferred Qualifications
  • Previous experience in the telecommunications industry or with high‑volume customer, network, usage, billing, CDR, XDR, or event data.
  • Hands‑on experience with Apache Iceberg, Apache Hudi, or Delta Lake.
  • Experience with Spark, Trino, Presto, StarRocks, ClickHouse, Hive, or similar analytical technologies.
  • Experience with open‑source data platform components such as MinIO, OpenMetadata, Apache Superset, dbt, or MLflow.
  • Experience deploying or operating data workloads on Kubernetes.
  • Familiarity with Docker, Helm, GitOps, and cloud‑native platform practices.
  • Experience with AWS, Microsoft Azure, Google Cloud Platform, or private cloud environments.
  • Knowledge of data governance, metadata management, data lineage, and data security concepts.
  • Experience with change data capture technologies such as Debezium or similar tools.
  • Familiarity with telecom data domains and large‑scale streaming architectures.
What We Value?
  • An engineering mindset focused on maintainability, scalability, and operational reliability.
  • Genuine interest in open‑source technologies and modern data architectures.
  • Ability to take ownership and independently drive technical tasks.
  • Willingness to challenge existing approaches and propose better solutions.
  • Clear communication and effective collaboration with both technical and business stakeholders.
  • A pragmatic approach that balances architectural quality with project delivery.
  • Gain hands‑on experience with modern big data and lakehouse technologies.
  • Contribute to architectural and technology decisions rather than only implementing predefined tasks.
  • Work with experienced data engineers, architects, and analytics professionals.
  • Build solutions using open‑source, cloud‑native, and vendor‑independent technologies.
  • Join an agile environment where technical ownership and measurable impact are highly valued.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer: Lakehouse & Real-Time Pipelines
Senior Data Engineer: Lakehouse & Real-Time Pipelines

Datazza • Fatih

On-site
TRY 500,000 - 900,000
Senior Big Data Engineer (Scala/Python + Spark)
Senior Big Data Engineer (Scala/Python + Spark)

Grid Dynamics • Fatih

On-site
TRY 1,072,000 - 1,501,000
Opportunity to work on bleeding-edge projects
Competitive salary
Flexible schedule
+1
Senior Big Data Engineer (Scala + Spark)
Senior Big Data Engineer (Scala + Spark)

Grid Dynamics • Fatih

On-site
TRY 1,269,000 - 2,116,000
Opportunity to work on bleeding-edge projects
Competitive salary
Flexible schedule
+1
Senior Data Engineer
Senior Data Engineer

Eti • Fatih

On-site
TRY 600,000 - 1,200,000
Senior Data Software Engineer with Databricks
Senior Data Software Engineer with Databricks

EPAM Systems • Turkey

On-site
TRY 2,915,000 - 4,373,000
Private health insurance
Mentoring programs
Learning platforms
+2
Lead Data Software Engineer with Databricks, Apache Kafka, Apache Spark, Kubernetes
Lead Data Software Engineer with Databricks, Apache Kafka, Apache Spark, Kubernetes

EPAM Systems • Turkey

On-site
TRY 5,348,000 - 8,751,000
Private health insurance
Learning & development
Mentoring programs
+1
Senior Data Engineer with Terraform
Senior Data Engineer with Terraform

EPAM Systems • Turkey

On-site
TRY 4,375,000 - 7,292,000
Private health insurance
Learning & development program
Mentoring programs
+2
Lead Data Engineer with Terraform
Lead Data Engineer with Terraform

EPAM Systems • Turkey

On-site
TRY 4,375,000 - 7,292,000
Private health insurance
Learning platforms access
Mentoring programs
Data Architect
Data Architect

EPAM Systems • Turkey

On-site
TRY 7,292,000 - 10,209,000
Continuous upskilling
Private health insurance
English courses
+3
Data Engineer
Data Engineer

Cypher Games • Fatih

On-site
TRY 600,000 - 900,000