Lead Data Platform Engineer

SRA Group

Toronto

Hybrid

CAD 110,000 - 170,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

SRA Group in Downtown Toronto, Canada, seeks a senior Data Engineer to architect and lead scalable data pipelines and analytics platforms. Hybrid role requiring at least 3 days in-office weekly with cross-functional collaboration.

You will design, build, and maintain pipelines on Hadoop/Databricks, lead data governance, and mentor engineers while delivering high-impact data products for global users.

Qualifications

  • 8+ years of experience in data engineering, big data analytics, or enterprise data platforms, including 2+ years in a lead or technical leadership role.
  • Strong Python experience (Pandas, NumPy, PySpark) with hands-on Impala.
  • Proven Hadoop-based data extraction, transformation, and processing experience.
  • Strong SQL skills with relational and distributed data stores.
  • Experience with cloud data platforms (Azure/AWS), Databricks, and Snowflake.
  • Hands-on ETL/ELT and data integration tooling (Airflow, NiFi, Azure Data Factory).
  • Experience in data modelling, querying, data mining, and reporting over large data volumes.
  • Exposure to machine learning concepts and AI data workflows is a plus.
  • Experience implementing CI/CD and DevOps practices for data engineering.

Responsibilities

  • Lead ingestion, transformation, aggregation, and processing of large-scale datasets for analytics and consumption.
  • Design, build, and maintain scalable data pipelines across Hadoop/Databricks and enterprise platforms with high data quality and reliability.
  • Drive data unification by integrating diverse data sources into a governed analytics foundation.
  • Collaborate with product managers, data science, platform strategy, and tech teams to translate requirements into scalable solutions.
  • Act as a technical bridge between business, analytics, and engineering, communicating architecture decisions and trade-offs.
  • Enable alignment across stakeholders to tie data solutions to business outcomes.
  • Identify innovation opportunities and deliver proofs of concept, prototypes, and pilots.
  • Provide technical leadership and mentorship to data engineers and analysts, promoting best practices in data modelling and pipeline design.
  • Promote data governance, quality, performance optimization, and maintainability.

Skills

Python
Pandas
NumPy
PySpark
Impala
Hadoop
SQL
CI/CD
DevOps
Data modelling

Tools

Apache Airflow
Apache NiFi
Azure Data Factory
Databricks
Snowflake
Azure
AWS

Job description

Location – Downtown Toronto (hybrid - minimum 3 days in a week)
Duration: 6 months with possible extensions
Key Responsibilities:
  • Lead the ingestion, transformation, aggregation, and processing of large scale datasets to enable advanced analytics and downstream consumption.
  • Design, build, and maintain robust, scalable data pipelines across Hadoop/Databricks and enterprise data platforms, ensuring high standards of data quality, reliability, performance, and availability.
  • Drive data unification initiatives, integrating multiple structured and semi structured data sources into a cohesive, governed analytical foundation.
Advanced Analytics Enablement
  • Manipulate and analyse high volume, high velocity, and high dimensional datasets using modern big data framework and/or Cloud native applications
  • Analyse large volumes of transactional and product data to produce insights and actionable recommendations that support business growth and value realisation.
  • Apply metrics, measurement frameworks, and benchmarking techniques to evaluate solution effectiveness and drive continuous improvement.
Cross Functional Collaboration
  • Partner with Product Managers, Data Science, Platform Strategy, and Technology teams to understand analytical and data requirements and translate them into scalable engineering solutions.
  • Act as a technical bridge between business, analytical, and engineering teams, clearly articulating architecture decisions, trade offs, and implementation approaches.
  • Enable alignment across stakeholders to ensure data solutions are directly tied to business and customer outcomes.
  • Identify innovation opportunities and deliver proofs of concept, prototypes, and pilot solutions aligned to near term and future business needs.
  • Integrate new and emerging data assets that enhance existing platforms, products, and services, strengthening overall value propositions.
  • Gather and synthesise feedback from clients, product, engineering, and sales teams to inform new solutions and product enhancements.
Technical Leadership & Mentorship
  • Provide technical leadership, guidance, and mentorship to data engineers and analysts, setting standards for engineering quality, scalability, performance, and maintainability.
  • Promote best practices in data modelling, pipeline design, performance optimisation, and data governance.
  • Influence engineering standards, architectural consistency, and long term platform sustainability.
All About You
Technical Skills & Experience
  • Strong proficiency in Python, including Pandas, NumPy, PySpark, with hands on experience using Impala.
  • Proven experience working on Hadoop based platforms, performing large scale data extraction, transformation, and processing.
  • Strong SQL skills and experience working with both relational and distributed data stores.
  • Experience with enterprise data platforms and business intelligence ecosystems.
  • Hands on experience with ETL / ELT and data integration tools, such as Apache Airflow, Apache NiFi, Azure Data Factory.
  • Experience in data modelling, querying, data mining, and reporting over large volumes of granular data.
  • Exposure to machine learning concepts and analytical techniques used in advanced data solutions and Feature calculations and Model serving is a big plus.
  • 8+ years of experience in data engineering, big data analytics, or enterprise data platforms, including 2+ years in a lead or technical leadership role.
  • Experience working with cloud based data platforms (Azure/AWS, Databricks/Snowflake), including data lakes, distributed compute, and storage services.
  • Experience implementing CI/CD pipelines and DevOps practices for data engineering workflows.
GenAI / LLM Skills (Preferred)
  • Experience enabling GenAI/AI products through scalable, reliable data ingestion and transformation pipelines (batch and streaming).
  • Exposure to unstructured and semi-structured data processing (documents/logs/text) and building curated datasets for downstream consumption.
  • Strong understanding of data governance, privacy, and security requirements when using enterprise data with AI (PII handling, access control, auditability).
  • Familiarity with operationalizing AI data workflows (monitoring, data quality checks, reproducibility, and cost-aware scaling in cloud environments).
Analytical & Business Acumen
  • Strong experience collecting, standardising, and summarising diverse datasets while identifying patterns, inconsistencies, and data quality issues.
  • Solid understanding of how analytics, metrics, and visualisation support business decision making.
  • Ability to comprehend complex operational systems and deliver scalable analytics and information products to a global user base.
Ways of Working
  • Comfortable operating in a fast paced, delivery driven environment, both as a hands on contributor and a technical leader.
  • Ability to move seamlessly between business, analytical, and technical contexts, communicating clearly with diverse audiences.
  • Demonstrates Client’s DQ values, with a collaborative, inclusive, and customer centric mindset.

Note: Skills which is highlighted are Mandatory for this position.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Data Engineer / Data Platform Lead
Lead Data Engineer / Data Platform Lead

Princeton IT Services, Inc • Toronto

Hybrid
CAD 120,000 - 180,000
Lead Data Scientist
Lead Data Scientist

Princeton IT Services, Inc • Toronto

Hybrid
CAD 140,000 - 190,000
Lead Data Engineer – Databricks
Lead Data Engineer – Databricks

TEEMA • Toronto

Hybrid
CAD 135,000 - 175,000
Hybrid work (3 days onsite)
40 hours per week
Team Lead, Data Engineering - Databricks to lead and mentor a team of data engineers, conducting code reviews, design reviews, and knowledge-sharing sessions across multiple locations
Team Lead, Data Engineering - Databricks to lead and mentor a team of data engineers, conducting code reviews, design reviews, and knowledge-sharing sessions across multiple locations

S.I. Systems Ltd. • Toronto

Hybrid
CAD 135,000 - 175,000
Data & Analytics lead
Data & Analytics lead

Altis Technology • Markham

Hybrid
CAD 100,000 - 125,000
Sr Platform Lead – Full Stack
Sr Platform Lead – Full Stack

Highbrow LLC • Toronto

Hybrid
CAD 120,000 - 150,000
Data Science Lead
Data Science Lead

Aarorn Technologies Inc • Vaughan

On-site
CAD 100,000 - 130,000
Lead Data Solution Engineer-6
Lead Data Solution Engineer-6

REALIGN LLC • Toronto

On-site
CAD 120,000 - 160,000
OPEN: Data Engineer
OPEN: Data Engineer

Cpus Engineering Staffing Solutions Inc. • Oshawa

On-site
CAD 80,000 - 100,000
Cloud Data Engineer | Toronto, ON | Fulltime FTE
Cloud Data Engineer | Toronto, ON | Fulltime FTE

Acestack • Toronto

Hybrid
CAD 90,000 - 140,000