MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Natural Language Processing Techniques
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

3,084 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
3,084 works in the cohort · of 4,299,418page 26 of 62

Labels cover 10 of 3,084 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 3,084 of 3,084 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

affunlabeled
RankLLM: A Python Package for Reranking with LLMs
Sahel Sharifymoghaddam, Ronak Pradeep, Andre Slavescu, Ryan Nguyen, Andrew Xu, Zijian Chen +4 more
2025· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
5
citations
aboutno affno abstractunlabeled
Proceedings of CoNLL-2003, Edmonton, Canada
Walter Daelemans, Michael Osborne
2003· article· ca· Data Archiving and Networked Services (DANS)· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
affunlabeled
Knowledge patterns in corpora
Elizabeth Marshman
2022· book-chapter· en· Terminology and lexicography research and practice· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
venueno affno abstractunlabeled
SYDNEY CMCRC at TAC 2013.
Glen Pink, Andrew Naoum, Will Radford, Will Cannings, Joel Nothman, Daniel Tse +1 more
2013· article· en· Theory and applications of categories· Computer Science
machine prediction:candidate · insufficient_payloadconsensus · none
4
citations
aboutno affunlabeled
Chaos out of Order
Raluca Tanasescu
2019· article· en· DOAJ (DOAJ: Directory of Open Access Journals)· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
affunlabeled
Experiments for HARD and Enterprise Tracks.
Olga Vechtomova, Maheedhar Kolla, Murat Karamuftuoglu
2005· article· en· Text REtrieval Conference· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
affunlabeled
Utilizing Extra-Sentential Context for Parsing
Jackie Chi Kit Cheung, Gerald Penn
2010· article· en· Empirical Methods in Natural Language Processing· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
afffundunlabeled
Compositional Generalization in Dependency Parsing
Emily J. Goodwin, Siva Reddy, Timothy J. O’Donnell, Dzmitry Bahdanau
2022· article· en· Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
affunlabeled
Swordfish
Chris Jordan, John Healy, Vlado Kešelj
2006· article· en· Computer Science
machine prediction:candidate · insufficient_payloadconsensus · none
4
citations
affno abstractunlabeled
Lexical Profiling of Environmental Corpora
Patrick Drouin, Marie-Claude L’Homme, Benoît Robichaud
2018· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
affno abstractunlabeled
Rhetorical Figure Annotation with XML.
Sebastian Ruan, Chrysanne Di Marco, Randy Allen Harris
2016· article· en· International Joint Conference on Artificial Intelligence· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations
affunlabeled
Handling Pronouns Intelligently
Anna Maria Di Sciullo
2005· article· en· New Trends in Software Methodologies, Tools and Techniques· Computer Science
machine prediction:candidate · noneconsensus · none
4
citations

How this was built: Screen · Findings · About