Senior Big Data Engineer: Spark, Hadoop & API Pipelines

Galent

Toronto

On-site

CAD 120,000 - 180,000

Full time

34 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Galent in Toronto is seeking a Senior Backend Developer to design, build, and maintain scalable data pipelines and backend systems in an enterprise environment. You will lead Spark-Scala projects on Hadoop/CDP clusters, optimize ETL pipelines, and migrate Spark 2 apps to Spark 3 while ensuring data governance and security.

The role requires 5+ years of experience in big data engineering, API integration, and AI-assisted development, with hands-on work on Spark, Hive, HDFS, and CI/CD tooling.

Qualifications

  • 5+ years of backend or data engineering experience in enterprise environments.
  • Hands-on with Spark, Hadoop, and data lake ecosystems.
  • Experience migrating Spark 2 to Spark 3 and building scalable ETL pipelines.

Responsibilities

  • Design and develop Spark-Scala apps for large-scale data processing on Hadoop/CDP clusters.
  • Build and optimize ETL/ELT pipelines using Spark DataFrames, Datasets and Spark SQL.
  • Tune Spark jobs for performance (partitioning, caching, broadcast joins, shuffle optimization).
  • Migrate Spark 2 apps to Spark 3 on Cloudera CDP platforms.
  • Work with Parquet, ORC, Avro on HDFS.
  • Write complex HiveQL / Spark SQL queries with window functions, CTEs, subqueries and aggregations.
  • Design and maintain Hive external/managed tables and partitioned datasets.
  • Optimize slow-running queries and resolve correlated subquery issues.
  • Work with HDFS encryption zones and data governance requirements.
  • Develop and maintain bash scripts for job orchestration and automation.
  • Handle error management, logging and alerting in shell scripts.
  • Manage HDFS operations (hdfs dfs commands), file transfers, and data validation.
  • Build scripts and pipelines to extract data from REST APIs using curl and Python.
  • Parse and process JSON API responses and load into HDFS/Hive.
  • Manage pagination, error handling and retry logic for API calls.
  • Work with enterprise API gateways and URL parameter construction.
  • Leverage GitHub Copilot / AI coding assistants to accelerate development.
  • Use AI tools for code review, SQL generation, script debugging and documentation.
  • Contribute to AI-assisted data quality and anomaly detection pipelines.
  • Explore and implement LLM-based automation for repetitive data engineering tasks.

Skills

Big data engineering
API integration
AI-assisted development

Tools

Spark
Hadoop
Hive
HDFS
CDP
Spark SQL
Parquet/ORC/Avro

Job description

Galent in Toronto is seeking a Senior Backend Developer to design, build, and maintain scalable data pipelines and backend systems in an enterprise environment. You will lead Spark-Scala projects on Hadoop/CDP clusters, optimize ETL pipelines, and migrate Spark 2 apps to Spark 3 while ensuring data governance and security.

The role requires 5+ years of experience in big data engineering, API integration, and AI-assisted development, with hands-on work on Spark, Hive, HDFS, and CI/CD tooling.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Java Spark Engineer, Data Pipelines & AI-Enabled Dev
Senior Java Spark Engineer, Data Pipelines & AI-Enabled Dev

Capgemini • Montreal (administrative region)

On-site
CAD 70,000 - 95,000
Big Data Developer
Big Data Developer

Galent • Toronto

On-site
CAD 120,000 - 180,000
Senior Data Engineer - Big Data & Analytics (AWS/Spark)
Senior Data Engineer - Big Data & Analytics (AWS/Spark)

Capgemini • Mississauga

On-site
CAD 110,000 - 170,000
Senior Data Engineer (Python, Spark, Snowflake) - Toronto, ON
Senior Data Engineer (Python, Spark, Snowflake) - Toronto, ON

Techedin • Toronto

Hybrid
Senior Java Spark Engineer — Big Data & AI Tools
Senior Java Spark Engineer — Big Data & AI Tools

Capgemini • Montreal (administrative region)

On-site
CAD 38,000 - 59,000
Medical benefits
Dental benefits
Vision benefits
+1
Spark & Scala Developer with Java Expertise (Big Data)
Spark & Scala Developer with Java Expertise (Big Data)

Confiar Services • Toronto

On-site
CAD 110,000 - 150,000
Senior Data Engineer
Senior Data Engineer

Techedin • Toronto

Hybrid
Senior Data & Backend Engineer (ETL & API)
Senior Data & Backend Engineer (ETL & API)

Autodesk • Toronto

On-site
CAD 107,000 - 157,000
Developer (Spark and Scala Exp)
Developer (Spark and Scala Exp)

Confiar Services • Toronto

On-site
CAD 110,000 - 150,000
Big Data Systems Engineer
Big Data Systems Engineer

Pyramid Consulting, Inc • Toronto

Hybrid
CAD 77,000 - 79,000