Data Warehouse Engineer

Mastronardi Produce Limited

Kingsville

On-site

CAD 85,000 - 95,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mastronardi Produce Limited is seeking a Data Warehouse Engineer to enhance and maintain our One Data Platform (ODP) built on Microsoft Fabric. You will profile data, design scalable data models, and optimize pipelines using PySpark and SQL in a Bronze/Silver/Gold architecture.

The role requires strong Python and Spark, deep SQL expertise, and the ability to guide other data engineers while collaborating with Finance, Supply Chain, and Reports teams.

Qualifications

  • Bachelor’s degree in computer science, information systems, or a related field.
  • Strong hands-on experience with Microsoft Fabric (Lakehouses, Data Pipelines/Dataflows, Notebooks) and Medallion (Bronze/Silver/Gold) architecture.
  • Expert-level SQL with complex joins, window functions, and performance tuning, plus ability to design SQL-first data models.
  • Experience profiling data sources and selecting ingestion/refresh strategies; knowledge of data governance is a plus.

Responsibilities

  • Profile and analyze data to translate business and reporting requirements into performant, well-modeled Gold-layer datasets.
  • Onboard new data sources and determine ingestion methods and refresh approaches (full load, incremental, CDC, or near-real-time).
  • Design, build, and optimize data pipelines, aiming to reduce run times and increase reliability at scale.
  • Own root-cause analysis and resolution of data engineering issues, collaborating with source system owners.
  • Build and evolve a flexible data foundation enabling end users to write SQL and build reports against governed models.
  • Collaborate with report developers to ensure accurate, well-modeled data that meets requirements.
  • Partner with Finance, Supply Chain, Sales, Operations, and Logistics to translate business needs into scalable data models.
  • Review others’ work, provide guidance on approach and best practices to raise engineering quality.
  • Think holistically across projects to scale the data platform as one cohesive solution.
  • Implement data governance, security, and clear documentation for data models and pipelines.
  • Operate as a team player in a fast-paced, evolving environment.

Skills

SQL
Spark
Python
Data modeling
Data profiling
Pipeline optimization
Troubleshooting
Stakeholder collaboration
Cross-functional partnership
Peer guidance

Education

Bachelor’s degree in computer science, information systems, or a related field

Tools

Microsoft Fabric
Medallion architecture
Synapse
Amazon S3
Azure Data Lake Storage
Google Cloud Storage
Git

Job description

Description

Mastronardi Produce pioneered the commercial greenhouse industry in North America, and we’re now the leading greenhouse vegetable company on the continent. Our award-winning, flavorful produce is packed under the SUNSET® brand and is available at leading grocery retailers across North America. Family owned for over 70 years, we pride ourselves on having the most flavorful products and the best people in the industry. We are constantly pushing boundaries to be a leader in fresh produce innovation. We seek individuals that demonstrate our PRIDE values (Passion, Respect, Innovation, Drive, Excellence) to help us fulfill our mission to inspire healthy living through WOW flavor experiences.

Our head office in Kingsville, ON is currently seeking a Data Warehouse Engineer to join our team! As a Data Warehouse Engineer, you will help enhance and maintain Mastronardi Produce’s One Data Platform (ODP), our enterprise data foundation built on Microsoft Fabric using Medallion (Bronze/Silver/Gold) architecture. Most development is done in PySpark notebooks, so strong coding skills are essential. You will profile and analyze data to meet stakeholder requirements, build and optimize pipelines, and design data models that are both technically sound and easy for end users to consume or to build their own reports. This role requires strong SQL, Spark/Python, and data modeling craft, sound engineering judgment across the full pipeline lifecycle, and the ability to guide and review the work of other data engineers.

Values:

To perform the job successfully, the incumbent’s behavior must be consistent with the PRIDE values expected of all Mastronardi Produce employees: be Passionate; have Respect; be Innovative; be Driven and strive for Excellence.

Primary Responsibilities:

  • Data modeling & stakeholder requirements — Profile and analyze data to translate business and reporting requirements into performant, well-modeled Gold-layer datasets.
  • New data source onboarding — Profile new data sources and determine the appropriate ingestion method and refresh approach (full load, incremental, CDC, or near-real-time) based on source system constraints, data volume, and business need.
  • Pipeline development & optimization — Design, build, and continuously optimize data pipelines, with a focus on reducing pipeline run times and improving reliability at scale.
  • Troubleshooting & root cause resolution — Own the investigation, root-causing, and resolution of data engineering-related pipeline and data issues, partnering with source system owners where needed.
  • Self-service data foundation — Build and evolve a data foundation that is flexible and easy for end users to work with directly — enabling them to write their own SQL and build their own reports against governed, well-documented Gold-layer models.
  • Partnership with report developers — Work closely with report developers throughout the build process to ensure requirements are met and that final data is accurate, well-modeled, and performant.
  • Partnership with business stakeholders — Partner with stakeholders across Finance, Supply Chain, Sales, Operations, and Logistics to understand data requirements and business logic, and translate them into scalable data models.
  • Guidance to other engineers — Review the work of other data engineers and provide guidance on approach, technique, and best practices, helping raise the overall quality and consistency of engineering across ODP.
  • Holistic platform thinking — Think across projects and domains — rather than in isolation — to build and enhance a data foundation that scales as one cohesive platform instead of disconnected, duplicative solutions.
  • Governance & documentation — Implement data governance, security, and access control practices consistent with ODP standards, and maintain clear documentation for data models, pipelines, and workflows.
  • Team orientation — Operate as a team player who executes on what’s best for the business, collaborating effectively with engineers, report developers, and stakeholders in a fast-paced, evolving environment.
  • Bachelor’s degree in computer science, information systems, or a related field.
  • MS Fabric Platform: Strong Hands-on experience with Microsoft Fabric (Lakehouses, Data Pipelines/Dataflows, Notebooks) and strong working knowledge of Medallion architecture (Bronze/Silver/Gold layer design).
  • Other Cloud Data Platforms: Good hands-on experience with Synapse, Amazon, S3, Azure Data Lake Storage or Google Cloud Storage.
  • SQL Expertise: Expert-level SQL, including complex joins, window functions, and query performance tuning, with the ability to design SQL-first data models that business users and report developers can query directly and confidently
  • Data Modeling: Strong data modeling skills, including dimensional (star schema) modeling, grain definition, and slowly changing dimensions, with an emphasis on designing schemas that are both engineering-sound and easy for end users to navigate.
  • Spark & Python: Expert-level PySpark and Python, as most Fabric development is done in Spark notebooks — including building and optimizing transformations, and writing clean, maintainable, production-grade notebook code.
  • Data Source Profiling: Demonstrated ability to profile unfamiliar data sources and recommend the right ingestion and refresh strategy based on source constraints, volume, and business need.
  • Pipeline Optimization: Experience identifying performance bottlenecks and reducing pipeline run times in a production data engineering environment.
  • Troubleshooting: Strong root-cause analysis and troubleshooting skills for data quality and pipeline issues.
  • Stakeholder & Cross-functional Partnership: Experience gathering requirements directly from business stakeholders and partnering with report developers to validate data accuracy and performance.
  • Peer Guidance: Experience reviewing others’ technical work and providing constructive, actionable guidance, even without formal management authority.

Specific Knowledge, Skills and Abilities Required

  • Broader platform exposure: Experience with other big data or cloud platforms (e.g., Hadoop, Databricks, AWS, Azure, or GCP) is a plus, though the core platform for this role is Microsoft Fabric.
  • AI-assisted Development: Experience using AI coding assistants and LLM-based tools (e.g., Claude, GitHub Copilot) to accelerate data engineering work, writing and reviewing SQL/PySpark, debugging pipelines, and generating documentation while retaining full ownership of correctness, performance, and quality.
  • Real-time Processing: Familiarity with real-time/streaming data processing (e.g., Kafka, Flink, Fabric Eventstream) is a plus. Note: near-real-time pipelines have not yet been built on ODP due to source system constraints and competing priorities, so this is a forward-looking skill rather than a day-one requirement.
  • Governance Tooling: Familiarity with data governance frameworks, data quality management, and metadata/catalog tools.
  • Tooling & Process: Experience with version control (GIT) and agile development practices.
  • Certifications in relevant technologies (e.g., Fabric Analytics Engineer. AWS Certified Big Data Specialty) are a plus.

Working Conditions:

  • Typical office environment.

Please note: Mastronardi Produce has accommodation processes and policies in place and provides accommodation for employees with disabilities. If you require a specific accommodation because of a disability or documented medical need, please contact the Human Resource office so that arrangements can be made for the appropriate accommodation to be put into place.

Salary is $85k/yr-$95k/yr CAD

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Analytics Engineer
Data Analytics Engineer

Mastronardi Produce Limited • Kingsville

On-site
CAD 85,000 - 95,000
Purchasing Analyst
Purchasing Analyst

Mastronardi Produce Limited • Kingsville

On-site
CAD 60,000
Cloud Developer
Cloud Developer

Mastronardi Produce • Kingsville

On-site
CAD 100,000 - 130,000
Buyer
Buyer

Mastronardi Produce • Kingsville

On-site
CAD 65,000 - 70,000
Maintenance Manager
Maintenance Manager

Mastronardi Produce • Sombra

On-site
CAD 72,000 - 88,000
Accommodation policy for disabilities
AI-powered hiring
Production Technician
Production Technician

Mastronardi Produce Limited • Delta

On-site
CAD 42,000 - 54,000
Production Technician
Production Technician

Socket.dev • Delta

On-site
CAD 29,000 - 30,000
Staff Data Engineer
Staff Data Engineer

Loblaw Digital • Toronto

On-site
CAD 198,000 - 268,000
Manager, Maintenance
Manager, Maintenance

Envirofresh-Produce-Inc. • Sombra

On-site
CAD 68,000 - 92,000
Data Engineer – Microsoft Fabric
Data Engineer – Microsoft Fabric

Systematix • Vaughan

On-site
CAD 103,000 - 117,000