Paires

Data Engineer

Paires

Remote · Full Time

Be the first to apply

Experience
3–8 yrs
Salary
CAD 150,000 – CAD 250,000 / year
Openings
1
Posted
10 hours ago
Work mode
Work from home
Resume
Required to apply

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About Paires and Role Overview

Paires offers a platform where founders connect with the right investors from a global network, facilitating warm outreach that converts into meetings. The team is small, senior, profitable, and self-funded, delivering rapidly in a flat structure.

This first Data Engineer position is crucial for owning and managing the primary database that powers all agents and outreach activities. The database is a complex knowledge graph combining companies, investors, funding rounds, news, and the underlying communication contexts such as emails and call transcripts.

Key Responsibilities

  • Design, scale, and maintain the main database architecture built on Postgres and Supabase with hybrid search capabilities, including schema design, data modeling, performance optimization, and growth management without limits.
  • Ensure high data quality through validation gates, deduplication, entity resolution, provenance tracking, and vigilant monitoring of vendor and third-party data.
  • Manage the communications layer by storing and linking raw emails and call transcripts to corresponding people and companies, enabling effective searchability.
  • Develop and optimize ingestion and enrichment pipelines to handle funding rounds, market news, and contact/company research with a focus on cost-efficiency and data freshness.
  • Maintain the knowledge graph structure with entities and relationships representing companies, investors, rounds, and news, ensuring every fact has provenance.
  • Create and uphold a unified clean data layer that serves all campaigns, agents, and product features consistently.

Candidate Profile

  • Experience owning and maintaining databases related to companies, people, deals, or communications, particularly those serving as a CRM or market intelligence knowledge source for live products and sales teams.
  • Proficient in SQL and Python with hands-on experience in building robust data pipelines involving ingestion, transformation, deduplication, and enrichment.
  • Track record of identifying and mitigating bad data before it impacts business operations.
  • Strong schema and contract design skills to future-proof database queries and structure.
  • Expertise in entity-relationship modeling at scale involving complex connections like investors, companies, funding rounds, and people.
  • Comfortable working quickly with AI tools and accountable for delivering high-impact outcomes.
  • Experience owning CRM data, enrichment pipelines, and deduplication processes, possibly combined with growth or RevOps roles.
  • Bonus expertise includes working with pgvector and embeddings, knowledge graph modeling within relational databases, scaling funding-news ingestion, entity resolution, and self-built communication stores.

Work Conditions and Benefits

  • Fully remote work with asynchronous communication.
  • Work hours overlap partially with US Eastern time, with meetings primarily scheduled on Mondays and Thursdays to enable extended focus time.
  • Access to advanced AI tools such as Claude Code, Cursor, and leading AI models at the company's expense.
  • Collaboration with GTM leaders and founding engineers, contributing to foundational layers that support the entire product.

Compensation

The salary ranges between CA$150,000 and CA$250,000 annually.

Tools & software

Python required PostgreSQL required

How they work

Teamwork & Collaboration Attention to Detail Adaptability Accountability Results Orientation

Leave it if you'd like a reply — we won't use it for anything else.

Click to browse, drag & drop, or paste a screenshot

PNG, JPG, GIF, MP4, WebM, MOV · Max 20MB each · Up to 5 files

🤖
Online · instant AI help