- Experience
- 3–8 yrs
- Salary
- CAD 150,000 – CAD 250,000 / year
- Openings
- 1
- Posted
- 10 hours ago
- Work mode
- Work from home
- Resume
- Required to apply
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About Paires and Role Overview
Paires offers a platform where founders connect with the right investors from a global network, facilitating warm outreach that converts into meetings. The team is small, senior, profitable, and self-funded, delivering rapidly in a flat structure.
This first Data Engineer position is crucial for owning and managing the primary database that powers all agents and outreach activities. The database is a complex knowledge graph combining companies, investors, funding rounds, news, and the underlying communication contexts such as emails and call transcripts.
Key Responsibilities
- Design, scale, and maintain the main database architecture built on Postgres and Supabase with hybrid search capabilities, including schema design, data modeling, performance optimization, and growth management without limits.
- Ensure high data quality through validation gates, deduplication, entity resolution, provenance tracking, and vigilant monitoring of vendor and third-party data.
- Manage the communications layer by storing and linking raw emails and call transcripts to corresponding people and companies, enabling effective searchability.
- Develop and optimize ingestion and enrichment pipelines to handle funding rounds, market news, and contact/company research with a focus on cost-efficiency and data freshness.
- Maintain the knowledge graph structure with entities and relationships representing companies, investors, rounds, and news, ensuring every fact has provenance.
- Create and uphold a unified clean data layer that serves all campaigns, agents, and product features consistently.
Candidate Profile
- Experience owning and maintaining databases related to companies, people, deals, or communications, particularly those serving as a CRM or market intelligence knowledge source for live products and sales teams.
- Proficient in SQL and Python with hands-on experience in building robust data pipelines involving ingestion, transformation, deduplication, and enrichment.
- Track record of identifying and mitigating bad data before it impacts business operations.
- Strong schema and contract design skills to future-proof database queries and structure.
- Expertise in entity-relationship modeling at scale involving complex connections like investors, companies, funding rounds, and people.
- Comfortable working quickly with AI tools and accountable for delivering high-impact outcomes.
- Experience owning CRM data, enrichment pipelines, and deduplication processes, possibly combined with growth or RevOps roles.
- Bonus expertise includes working with pgvector and embeddings, knowledge graph modeling within relational databases, scaling funding-news ingestion, entity resolution, and self-built communication stores.
Work Conditions and Benefits
- Fully remote work with asynchronous communication.
- Work hours overlap partially with US Eastern time, with meetings primarily scheduled on Mondays and Thursdays to enable extended focus time.
- Access to advanced AI tools such as Claude Code, Cursor, and leading AI models at the company's expense.
- Collaboration with GTM leaders and founding engineers, contributing to foundational layers that support the entire product.
Compensation
The salary ranges between CA$150,000 and CA$250,000 annually.