MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Natural Language Processing Techniques
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

3,084 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
3,084 works in the cohort · of 4,299,418page 5 of 62

Labels cover 10 of 3,084 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 3,084 of 3,084 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

affunlabeled
Automatic Acquisition of Lexical Formality
Julian Brooke, Tong Wang, Graeme Hirst
2010· article· en· International Conference on Computational Linguistics· Computer Science
machine prediction:candidate · noneconsensus · none
56
citations
venueno affunlabeled
Evidence of Parallel Processing During Translation
Laura Winther Balling, Kristian Tangsgaard Hvelplund, Annette C. Sjørup
2014· article· en· Meta Journal des traducteurs· Computer Science
machine prediction:candidate · noneconsensus · none
56
citations
affunlabeled
UniMorph 3.0: Universal Morphology
Arya D. McCarthy, Christo Kirov, Matteo Grella, Amrit Nidhi, Patrick Xia, Kyle Gorman +15 more
2020· article· en· Minerva Access (University of Melbourne)· Computer Science
machine prediction:candidate · noneconsensus · none
54
citations
fundno affunlabeled
A Deep Architecture for Semantic Parsing
Edward Grefenstette, Phil Blunsom, Nando de Freitas, Karl Moritz Hermann
2014· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
54
citations
affunlabeled
<i>Yawat</i>
Ulrich Germann
2008· article· en· Computer Science
machine prediction:candidate · insufficient_payloadconsensus · none
54
citations
affunlabeled
Alignment-Based Discriminative String Similarity
Shane Bergsma, Grzegorz Kondrak
2007· article· en· Meeting of the Association for Computational Linguistics· Computer Science
machine prediction:candidate · noneconsensus · none
54
citations
aboutno affunlabeled
Native Languages of Alaska
Michael E. Krauss
2007· book-chapter· en· Computer Science
machine prediction:candidate · noneconsensus · none
54
citations
affunlabeled
Modeling Past and Future for Neural Machine Translation
Zaixiang Zheng, Hao Zhou, Shujian Huang, Lili Mou, Xinyu Dai, Jiajun Chen +1 more
2018· article· en· Transactions of the Association for Computational Linguistics· Computer Science
machine prediction:candidate · noneconsensus · none
54
citations
afffundunlabeled
HMM word recognition engine
D. Guillevic, C.Y. Suen
2002· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
52
citations
afffundunlabeled
Trans-scription as a social activity
Cécile B. Vigouroux
2007· article· en· Ethnography· Computer Science
machine prediction:candidate · noneconsensus · none
51
citations
affno abstractunlabeled
Natural Language Understanding
2013· book-chapter· en· Computer Science
machine prediction:candidate · noneconsensus · none
50
citations
affaboutunlabeled
Human interaction for high-quality machine translation
Francisco Casacuberta, Jorge Civera, Elsa Cubel, Antonio L. Lagarda, Guy Lapalme, Elliott Macklovitch +1 more
2009· article· en· Communications of the ACM· Computer Science
machine prediction:candidate · noneconsensus · none
49
citations

How this was built: Screen · Findings · About