Data Engineer

Paires

Canada

On-site

CAD 90,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote
Async work
US Eastern overlap hours
Batch meetings

Job summary

Paires is hiring its first Data Engineer to own the database that powers our agent outreach and knowledge graph. You will design, scale, and keep the data clean, serving as the memory of our live product.

This is a fully remote role with overlapping time with US Eastern hours. You will work on Postgres/Supabase, with pgvector and AI tooling, building end-to-end data pipelines and ensuring data quality across vendor and partner data sources.

Qualifications

  • Proven ability to own and scale a live database used by a product team.
  • Strong SQL and Python pipelines for ingestion, transformation, and deduplication.
  • Experience with data quality gates, entity resolution, and provenance.
  • Ability to design schemas and contracts for future queries.

Responsibilities

  • Own the database and its growth, with schema design, modeling, and performance tuning.
  • Build and maintain ingestion and enrichment pipelines for funding rounds, market news, and contacts.
  • Maintain data quality through validation gates and deduplication processes.
  • Develop and maintain the knowledge graph of companies, investors, deals, and people.

Skills

SQL
Python
Data modeling
Data quality
Schema design

Tools

Postgres
Supabase
pgvector
AI tooling

Job description

We are hiring our first Data Engineer to own the database our agents and outreach are built on.

Paires is where founders come to raise capital. We pair them with the right investors from a large, engaged global investor network, then run the warm outreach that turns into meetings. It is a two-sided platform, live with paying clients, profitable and self-funded, built by a small, senior, flat team that ships fast.

The role

Everything we do runs on one asset: a database of every company and investor out there, every funding round, the news that matters, and how they all connect - plus the raw context underneath: every email and call transcript, linked to the right people and companies. It is a knowledge graph and a memory in one. Our matching, our outreach, and our agents are built on top of it, and it grows faster than anyone can own it on the side. You become its owner. You design it, scale it, keep it clean, and turn it into the single source of truth that everything reads from. To be clear about the shape of this seat: it is not a reporting or analytics warehouse. It is the memory a live product thinks with, built for one reader above all: agents retrieving exactly the right fact at the right moment. One honest filter before you apply: if the database you are proudest of tracked shipments, sensors, factory lines, or compliance - however well you built it - that is a different seat. If it tracked companies, investors, deals, and the people and conversations around them, keep reading.

What you will own
  • The database itself: Postgres and Supabase with hybrid search, schema design, modeling, scaling, and performance as it grows without a ceiling. The agents that read it run on Pydantic AI and the Claude Agent SDK, on AWS. We are consolidating into pgvector, not buying a vector DB.

  • Data quality end to end: validation gates for vendor and third-party data, dedup, entity resolution, provenance, monitoring.

  • The communications layer: raw emails and call transcripts stored, linked to the right people and companies, and searchable.

  • Ingestion and enrichment pipelines: funding rounds, market news, and contact and company research at scale, engineered for cost and freshness.

  • The knowledge graph: companies, investors, funding rounds, and news as entities and relationships - node and edge tables in Postgres, provenance on every fact.

  • The unified data layer: one clean spine that every campaign, agent, and product feature reads from.

You are a fit if you
  • Have owned a database of companies, people, deals, or the communications between them - a CRM source of truth, a market or deal intelligence graph, an enrichment layer - that a live product, agents, or a sales team read from. Serving dashboards is a different job than this one.

  • Are strong in SQL and Python, with real pipeline work behind you: ingest, transform, dedup, enrich.

  • Have caught bad data before it hurt the business, and can tell us how.

  • Think in schemas and contracts, and design for the queries of a year from now.

  • Have modeled entities and relationships at scale - companies to investors to rounds to people - and kept the connections queryable as the sources multiplied.

  • Move fast with AI tooling and own outcomes.

  • You do not need the title. If you were the RevOps or growth person who owned the CRM data, the enrichment pipelines, and the dedup nobody else wanted - and you got real hands-on with AI - we want to hear from you.

  • Bonus: pgvector and embeddings, a knowledge graph you modeled in a relational database, funding-round or news ingestion at scale, entity resolution at scale, a raw communications store you built yourself.

What we offer
  • Fully remote and async.

  • Your day overlaps with US Eastern time for a few hours - not full US hours.

  • Meetings batch on Mondays and Thursdays, the rest is deep work.

  • The best AI tooling, paid (Claude Code, Cursor, top models).

  • You work alongside our GTM lead and our founding engineers, and your layer feeds everything they build.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Founding AI Engineer
Founding AI Engineer

Paires • Canada

On-site
CAD 120,000 - 190,000
Fully remote
Equity potential
GTM Engineer
GTM Engineer

Paires • Canada

On-site
CAD 168,208 - 252,313
Fully remote
Async work
Work with GTM lead
Head of Brand & Story
Head of Brand & Story

Paires • Canada

On-site
CAD 110,000 - 170,000
Fully remote work
Direct access to co-founders
Creative autonomy
+1
Senior Analytics Engineer
Senior Analytics Engineer

Passage • Toronto

On-site
CAD 100,000 - 150,000
Data Engineer
Data Engineer

Drive Capital • Toronto

Hybrid
CAD 80,000 - 100,000
Competitive salary
Professional development support
Dynamic work environment
Junior Product Manager
Junior Product Manager

Ampliwork, Inc • Montreal (administrative region)

On-site
CAD 80,000 - 110,000
Senior Product Manager
Senior Product Manager

Ampliwork • Montreal (administrative region)

On-site
CAD 120,000 - 180,000
Senior Product Manager
Senior Product Manager

Ampliwork, Inc • Montreal (administrative region)

On-site
CAD 120,000 - 180,000
Staff Database Engineer
Staff Database Engineer

Owner.com • Canada

On-site
USD 140,000 - 210,000
Investor Network Manager
Investor Network Manager

Paires • Canada

On-site
USD 110,000 - 180,000