MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Topic Modeling
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

2,769 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
2,769 works in the cohort · of 4,299,418page 24 of 56

Labels cover 6 of 2,769 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 2,769 of 2,769 works in this cohort. Predictions are machine_predicted_unvalidated teacher distillation outputs. Candidate is the union; consensus is the intersection.

affunlabeled
A word embedding trained on South African news data
Martin Mafunda, Maria Schuld, Kevin Durrheim, Sindisiwe Mazibuko
2022· article· en· The African Journal of Information and Communication (AJIC)· Computer Science
distilled prediction:candidate · noneconsensus · none
6
citations
affunlabeled
Large Language Model (LLM) Bias Index -- LLMBI
Abiodun Finbarrs Oketunji, Muhammad Anas, Deepthi Saina
2023· preprint· en· arXiv (Cornell University)· Computer Science
distilled prediction:candidate · metaepi_narrowconsensus · none
6
citations
afffundvenueunlabeled
Natural Language Processing for Virtual Reference Analysis
Ansh Sharma, Kathryn Barrett, Kirsta Stapelfeldt
2022· article· en· Evidence Based Library and Information Practice· Computer Science
distilled prediction:candidate · scholarly_communicationconsensus · none
6
citations
afffundunlabeled
Analogy Training Multilingual Encoders
Nicolas Garneau, Mareike Hartmann, Anders Sandholm, Sebastian Ruder, Ivan Vulić, Anders Søgaard
2021· article· en· Proceedings of the AAAI Conference on Artificial Intelligence· Computer Science
distilled prediction:candidate · noneconsensus · none
5
citations
affunlabeled
The impact of corpus size on question answering performance
Charles L. A. Clarke, Gordon V. Cormack, Michael Laszlo, Thomas R. Lynam, Egidio L. Terra
2002· article· en· Proceedings of the 25th annual international ACM SIGIR conference on Research and development in information retrieval - SIGIR '02· Computer Science
distilled prediction:candidate · noneconsensus · none
5
citations
affunlabeled
CBench
Abdelghny Orogat, Ahmed El-Roby
2021· article· en· Proceedings of the VLDB Endowment· Computer Science
distilled prediction:candidate · noneconsensus · none
5
citations
afffundgemma · no categorygpt · no categorymodels agree
Discourse relations in rationale‐containing text‐segments
Lu Xiao, Nadia Conroy
2017· article· en· Journal of the Association for Information Science and Technology· Computer Science
distilled prediction:candidate · noneconsensus · none
5
citations

How this was built: Screen · Findings · About