MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Topic Modeling
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

2,769 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
2,769 works in the cohort · of 4,299,418page 11 of 56

Labels cover 6 of 2,769 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 2,769 of 2,769 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

fundno affunlabeled
Conditional probing: measuring usable information beyond a baseline
John Hewitt, Kawin Ethayarajh, Percy Liang, Christopher D. Manning
2021· article· en· Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing· Computer Science
machine prediction:candidate · noneconsensus · none
24
citations
venueno affunlabeled
CUNY-UIUC-SRI TAC-KBP2011 Entity Linking System Description
Taylor Cassidy, Zheng Chen, Javier Artiles, Heng Ji, Hongbo Deng, Lev-Arie Ratinov +3 more
2011· article· en· Theory and applications of categories· Computer Science
machine prediction:candidate · noneconsensus · none
24
citations
afffundunlabeled
You Only Need Attention to Traverse Trees
Mahtab Ahmed, Muhammad Rifayat Samee, Robert E. Mercer
2019· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
23
citations
fundno affunlabeled
P-SIF: Document Embeddings Using Partition Averaging
Vivek Gupta, Ankit Saw, Pegah Nokhiz, Praneeth Netrapalli, Piyush Rai, Partha Talukdar
2020· article· en· Proceedings of the AAAI Conference on Artificial Intelligence· Computer Science
machine prediction:candidate · noneconsensus · none
23
citations
affno abstractunlabeled
Evaluating Corroborative Evidence
Douglas Walton, Chris Reed
2008· article· en· Argumentation· Computer Science
machine prediction:candidate · noneconsensus · none
22
citations
affunlabeled
Strategies of text retrieval: A criterion shift account.
Murray Singer, Nathalie Gagnon, Eric Richards
2002· article· en· Canadian Journal of Experimental Psychology/Revue canadienne de psychologie expérimentale· Computer Science
machine prediction:candidate · noneconsensus · none
22
citations
affunlabeled
A Survey of Conversational Search
Fengran Mo, Kelong Mao, Ziliang Zhao, Hongjin Qian, Haonan Chen, Yiruo Cheng +4 more
2025· article· en· ACM Transactions on Information Systems· Computer Science
machine prediction:candidate · noneconsensus · none
20
citations

How this was built: Screen · Findings · About