MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Natural Language Processing Techniques
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

3,084 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
3,084 works in the cohort · of 4,299,418page 15 of 62

Labels cover 10 of 3,084 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 3,084 of 3,084 works in this cohort. Predictions are machine_predicted_unvalidated teacher distillation outputs. Candidate is the union; consensus is the intersection.

affunlabeled
Resolving this-issue anaphora
Varada Kolhatkar, Graeme Hirst
2012· article· en· Computer Science
distilled prediction:candidate · noneconsensus · none
13
citations
affunlabeled
Machine Learning in Natural Language Processing
Marina Sokolova, Stan Śzpakowicz
2010· book-chapter· en· IGI Global eBooks· Computer Science
distilled prediction:candidate · metaepi_narrow+research_integrityconsensus · none
12
citations
venueno affno abstractunlabeled
Domination in lexicographic product digraphs.
Juan Liu, Xindong Zhang, Jixiang Meng
2011· article· en· Ars Combinatoria· Computer Science
distilled prediction:candidate · noneconsensus · none
12
citations
affunlabeled
Metaphor, Simulation, and Fictive Motion
Teenie Matlock
2017· book-chapter· en· Cambridge University Press eBooks· Computer Science
distilled prediction:candidate · metaepi_narrowconsensus · none
12
citations
affunlabeled
UofL
Yllias Chali, Shafiq Joty
2007· article· en· Computer Science
distilled prediction:candidate · noneconsensus · none
12
citations
venueno affno abstractunlabeled
TinkerBell: Cross-lingual Cold-Start Knowledge Base Construction.
Mohamed Al-Badrashiny, Jason Bolton, Arun Tejasvi Chaganty, Kevin B. Clark, Craig Harman, Lifu Huang +24 more
2017· article· en· Theory and applications of categories· Computer Science
distilled prediction:candidate · noneconsensus · none
12
citations
affunlabeled
Thematic indirect objects in French
Yves Roberge, Michelle Troberg
2007· article· en· Journal of French Language Studies· Computer Science
distilled prediction:candidate · noneconsensus · none
12
citations
venueno affunlabeled
Lexical based two-way RTE System at RTE-5.
Partha Pakray, Sivaji Bandyopadhyay, Alexander Gelbukh
2009· article· en· Theory and applications of categories· Computer Science
distilled prediction:candidate · noneconsensus · none
12
citations

How this was built: Screen · Findings · About