Data Engineer

Soroc Technology

Toronto

On-site

CAD 110,000 - 150,000

Full time

41 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Soroc Technology in Toronto is seeking an experienced Data Engineer to design, build, and support enterprise data platforms. You will develop scalable pipelines with PySpark/Spark, implement real-time processing with Kafka and NiFi, and optimize ingestion and transformations across large datasets.

The role requires strong SQL and Python, familiarity with Hadoop components (HDFS, Hive, YARN), and collaboration with architects and analytics teams to ensure data quality, governance, and reliable

Qualifications

  • Hands-on experience with PySpark/Spark stack and data ingestion pipelines.
  • Experience designing and optimizing batch and real-time data processing solutions.
  • Proficiency in SQL and Python for data extraction, transformation, and validation.
  • Familiarity with Hadoop ecosystem components (HDFS, Hive, YARN) and data warehousing concepts.

Responsibilities

  • Design, develop, and optimize scalable data pipelines using PySpark, Spark, Hadoop, and Apache NiFi.
  • Build and maintain batch and real-time data processing solutions.
  • Develop and support Kafka-based streaming applications and event-driven architectures.
  • Create and optimize ETL/ELT workflows for large-scale datasets.
  • Develop complex SQL queries for data extraction and validation.
  • Implement data ingestion from databases, APIs, files, and streaming sources.
  • Monitor and optimize Spark jobs and data pipelines.
  • Collaborate with architects and analytics teams to deliver high-quality data solutions.
  • Support platform upgrades, deployments, testing, and production releases.
  • Ensure data quality, governance, security, and operational excellence across data platforms.

Skills

PySpark
Apache Spark
Kafka
Hadoop
Apache NiFi
SQL
Python
Spark Streaming
Hive

Tools

Jenkins
Bitbucket
Git
JIRA
Confluence

Job description

An experienced Data Engineer with strong expertise in Big Data technologies to design, develop, and support enterprise-scale data platforms. The ideal candidate should possess hands‑on experience in PySpark, Apache Spark, Kafka, Hadoop ecosystem components, and Apache NiFi, with a strong understanding of data ingestion, transformation, and real‑time processing frameworks.

Key Responsibilities

  • Design, develop, and optimize scalable data pipelines using PySpark, Spark, Hadoop, and Apache NiFi.
  • Build and maintain batch and real‑time data processing solutions.
  • Develop and support Kafka‑based streaming applications and event‑driven architectures.
  • Create and optimize ETL/ELT workflows for large‑scale structured and unstructured datasets.
  • Develop complex SQL queries for data extraction, transformation, validation, and troubleshooting.
  • Implement data ingestion solutions from databases, APIs, files, and streaming sources.
  • Monitor, troubleshoot, and enhance the performance of Spark jobs and data pipelines.
  • Collaborate with architects, business analysts, and development teams to deliver high‑quality data solutions.
  • Support platform upgrades, deployments, testing, certification, and production releases.
  • Ensure data quality, governance, security, and operational excellence across data platforms.

Mandatory Skills

  • PySpark
  • Apache Spark (Spark SQL, DataFrames)
  • Hadoop Ecosystem (HDFS, Hive, YARN)
  • SQL
  • Python

Preferred Skills

  • Spark Streaming
  • Hive
  • Scala
  • Jenkins, Bitbucket, Git
  • JIRA, Confluence
  • Data Warehousing concepts and Dimensional Modeling
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer (Big Data)
Data Engineer (Big Data)

Pacer Group • Toronto

On-site
CAD 90,000 - 130,000
Data Engineer
Data Engineer

ALLTECH CONSULTING SVC INC • Mississauga

On-site
CAD 80,000 - 120,000
Data Engineering Lead and Data Engineers
Data Engineering Lead and Data Engineers

ALLTECH CONSULTING SVC INC • Ottawa

On-site
CAD 90,000 - 130,000
Data Engineer
Data Engineer

MPA Recruitment • Toronto

On-site
CAD 120,000 - 160,000
Data Engineer
Data Engineer

TalentOla • Mississauga

On-site
CAD 90,000 - 110,000
Data Engineer/Developer with Python and SQL
Data Engineer/Developer with Python and SQL

TechDoQuest • Mississauga

On-site
CAD 75,000 - 110,000
Sr Data Engineer
Sr Data Engineer

Tailored Brands, Inc. • Cambridge

On-site
CAD 120,000 - 180,000
Big Data Engineer
Big Data Engineer

KTek Resourcing • Mississauga

On-site
CAD 70,000 - 110,000
Senior Data Engineer
Senior Data Engineer

Capgemini • Mississauga

On-site
CAD 110,000 - 170,000
Big Data Developer – Scala/Spark, Java
Big Data Developer – Scala/Spark, Java

Astra-North Infoteck Inc. ~ Conquering today’s challenges, achieving tomorrow’s vision! • Toronto

On-site
CAD 90,000 - 130,000