Data Development Intern

Environics Analytics

Toronto

On-site

CAD 40,000 - 60,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Environics Analytics in Toronto invites a Data Development Intern to help build and modernize data infrastructure powering demographic and behavioral products. You will design automated pipelines in SQL and Python, support migration to Snowflake, and contribute to QC systems ensuring accurate outputs across geographies.

This hands-on role offers real ownership of production code and exposure to large-scale datasets from Statistics Canada and other sources, with collaboration across researchers

Qualifications

  • Enrolled in or recently completed a graduate program (Master's) in Computer Science, Data Science, Statistics, Geography, Engineering, or a related quantitative field.
  • Undergraduate candidates with strong relevant experience will also be considered.
  • Prior experience in data engineering, data analysis, or software development.
  • Comfort working with large-scale structured datasets (millions of rows across related tables).

Responsibilities

  • Design, build, and maintain automated data pipelines for ETL, modelling, and quality control across demographic data products.
  • Help migrate and refactor legacy workflows into SQL (T-SQL) and Python, improving scalability, maintainability, and version control.
  • Develop stored procedures, temp table-based workflows, and batch scripts to support large-scale data transformation.
  • Build automated QC checks and validation logic to catch anomalies and inter-vintage inconsistencies early in the pipeline.
  • Collaborate with data developers, Research Associates, and Technical Leads to translate data product methodology into reliable, repeatable code.
  • Present design approaches before building, validate results after, and participate in code reviews.
  • Use Azure DevOps and Git for version control and work item tracking; maintain documentation on SharePoint.
  • Investigate and prototype new tools, libraries, or pipeline architectures, including Snowflake-native approaches, that improve team efficiency or product quality.
  • Use AI coding tools (e.g., GitHub Copilot) as a core part of daily development to accelerate scripting, refactoring, and code review.
  • Apply AI-assisted approaches to documentation and QC, such as generating test cases, drafting validation logic, or summarizing pipeline behaviour.
  • Critically evaluate AI-generated code and output, verifying correctness and understanding the underlying SQL/Python well enough to own what ships.

Skills

SQL
Python
pandas
VS Code
Jupyter Notebook
SQL Server Management Studio
Git
Azure DevOps
GitHub Copilot

Education

Master's degree in Computer Science, Data Science, Statistics, Geography, Engineering
Undergraduate candidates with strong relevant experience

Tools

Snowflake
Airflow
Dask
ETL tools
APIs

Job description

Role Objective

The Data Development Intern plays a central role in building, maintaining, and modernizing the data infrastructure that powers Environics Analytics' core demographic and behavioural data products. The Data Development team works across a wide range of data sources, including Statistics Canada, IRCC, CRA, and third‑party survey data, applying rigorous ETL, quality control, and modeling pipelines to produce market‑ready outputs from national to small‑area geographies. You'll design and implement automated data pipelines in SQL and Python, support the migration of legacy workflows to modern architecture (including the team's move to Snowflake), and contribute to quality control systems that ensure the accuracy and consistency of our data products across vintages. This is a hands‑on role with real ownership of production code, and strong performers will be well positioned for a full‑time Data Engineer role on the team.

What You'll Do
  • Design, build, and maintain automated data pipelines for ETL, modelling, and quality control across demographic data products.
  • Help migrate and refactor legacy workflows into SQL (T‑SQL) and Python, improving scalability, maintainability, and version control.
  • Develop stored procedures, temp table‑based workflows, and batch scripts to support large‑scale data transformation.
  • Build automated QC checks and validation logic to catch anomalies and inter‑vintage inconsistencies early in the pipeline.
  • Collaborate with data developers, Research Associates, and Technical Leads to translate data product methodology into reliable, repeatable code.
  • Present design approaches before building, validate results after, and participate in code reviews.
  • Use Azure DevOps and Git for version control and work item tracking; maintain documentation on SharePoint.
  • Investigate and prototype new tools, libraries, or pipeline architectures, including Snowflake‑native approaches, that improve team efficiency or product quality.
  • Use AI coding tools (e.g., GitHub Copilot) as a core part of daily development to accelerate scripting, refactoring, and code review.
  • Apply AI‑assisted approaches to documentation and QC, such as generating test cases, drafting validation logic, or summarizing pipeline behaviour.
  • Critically evaluate AI‑generated code and output, verifying correctness and understanding the underlying SQL/Python well enough to own what ships.
What You'll Learn
  • Practical, production experience in data engineering, automation, and AI‑assisted development.
  • Exposure to large‑scale demographic, financial, and behavioural datasets.
  • Insight into the full product development lifecycle at a leading data and analytics firm.
  • Modern cloud data warehousing (Snowflake) alongside traditional SQL Server workflows.
  • Agile development, version control, and code review practices.
  • Best practices in quality control and data integrity at scale.
Qualifications
Education

Enrolled in or recently completed a graduate program (Master's) in Computer Science, Data Science, Statistics, Geography, Engineering, or a related quantitative field. Undergraduate candidates with strong relevant experience will also be considered.

Experience
  • Prior experience (coursework, research, co‑op, or work) in data engineering, data analysis, or software development.
  • Comfort working with large‑scale structured datasets (millions of rows across related tables).
Technical Skills
  • Strong SQL, including window functions and set‑based transformation logic; T‑SQL experience is a plus.
  • Proficiency in Python for data processing and automation, including pandas.
  • Experience building or contributing to multi‑step ETL pipelines.
  • Comfort working in VS Code, Jupyter Notebook, and/or SQL Server Management Studio.
  • Experience with Git and a willingness to learn Azure DevOps.
  • Comfort using AI coding tools (e.g., GitHub Copilot) as part of your regular workflow.
Bonus Skills
  • Familiarity with Snowflake or other cloud data warehousing.
  • Familiarity with ETL processes and APIs.
  • Exposure to geospatial data or Canadian census geographies (e.g., DA, CT, CSD, CMA).
  • Familiarity with dashboards, data visualization, or statistical concepts (imputation, aggregation, index construction).
  • Exposure to workflow orchestration tools (e.g., Airflow) or distributed computing (e.g., Dask).
Personal Attributes
  • Strong problem‑solving skills and eagerness to learn; comfortable identifying root causes and proposing systematic fixes.
  • Detail‑oriented, with good documentation and communication habits.
  • Collaborative and open to feedback; comfortable working in a multidisciplinary team of researchers and data professionals.
  • Able to clearly communicate technical findings to both technical and non‑technical stakeholders.
About Environics Analytics

Environics Analytics (EA) is a marketing services company that specializes in geodemographic‑based segmentation, site evaluation modelling, and custom analytics.

EA is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. If you require any accommodation to participate in the hiring process, please note the request in your application. We welcome people of all abilities.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

OPEN: Data Engineer
OPEN: Data Engineer

Cpus Engineering Staffing Solutions Inc. • Oshawa

On-site
CAD 80,000 - 100,000
Data Engineering Intern: Pipelines & Snowflake
Data Engineering Intern: Pipelines & Snowflake

Environics Analytics • Toronto

On-site
CAD 40,000 - 60,000
Staff Data Engineer
Staff Data Engineer

Loblaw Digital • Toronto

On-site
CAD 198,000 - 268,000
Data Engineer
Data Engineer

Financeit • Toronto

On-site
CAD 75,000 - 95,000
Hybrid workplace
Competitive salary with bonus
Comprehensive benefits
+3
Senior Integrations Developer / AI Data Integration Engineer
Senior Integrations Developer / AI Data Integration Engineer

Engineered Intelligence Inc. • Mississauga

Hybrid
CAD 90,000 - 120,000
Competitive compensation and benefits
Flexible hours
Autonomy & Growth opportunities
Data Engineer - Snowflake, dbtCore/Cloud, AWS
Data Engineer - Snowflake, dbtCore/Cloud, AWS

Aviva Canada • Markham

Hybrid
CAD 100,000 - 125,000
Hybrid flexible work model
Annual bonus eligibility
Health benefits
+2
Senior Data Engineer
Senior Data Engineer

Aarorn Technologies Inc • Markham

On-site
Senior Data Engineer
Senior Data Engineer

Socket.dev • Toronto

Hybrid
CAD 110,000 - 160,000
Flexible work arrangements
Generous vacation policy
Customizable benefits
+1
Senior Data Engineer
Senior Data Engineer

BuzzClan • Edmonton

On-site
CAD 100,000 - 150,000
Senior Data Engineer
Senior Data Engineer

Cpus Engineering Staffing Solutions Inc. • Pickering

On-site
CAD 90,000 - 120,000