MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
→
Sort
Language
Type
Field
Venue
arXiv (Cornell University)
Topic
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

12,857 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
12,857 works in the cohort · of 4,299,418page 52 of 258

Labels cover 18 of 12,857 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 12,857 of 12,857 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

affunlabeled
Towards a General-Purpose Linguistic Annotation Backend
Graham Neubig, Patrick Littell, Chian-Yu Chen, Jean Lee, Zirui Li, Yu-Hsiang Lin +1 more
2018· preprint· en· arXiv (Cornell University)· Computer Science
machine prediction:candidate · noneconsensus · none
3
citations
fundno affunlabeled
Information criteria for non-normalized models
Takeru Matsuda, Masatoshi Uehara, Aapo Hyvärinen
2019· preprint· en· arXiv (Cornell University)· Mathematics
machine prediction:candidate · noneconsensus · none
3
citations
affunlabeled
Super-Acceleration with Cyclical Step-sizes
Baptiste Goujaud, Damien Scieur, Aymeric Dieuleveut, Adrien Taylor, Fabián Pedregosa
2021· preprint· en· arXiv (Cornell University)· Engineering
machine prediction:candidate · noneconsensus · none
3
citations
affunlabeled
XORRO: Rapid Paired-End Read Overlapper
Russell J. Dickson, Gregory B. Gloor
2013· preprint· en· arXiv (Cornell University)· Biochemistry, Genetics and Molecular Biology
machine prediction:candidate · noneconsensus · none
3
citations
affunlabeled
Filtering Variational Objectives
Chris J. Maddison, Dieterich Lawson, George Tucker, Nicolas Heess, Mohammad Norouzi, Andriy Mnih +2 more
2017· preprint· en· arXiv (Cornell University)· Computer Science
machine prediction:candidate · noneconsensus · none
3
citations
fundno affunlabeled
Uniform generation of random regular graphs
Pu Gao, Nicholas Wormald
2015· preprint· en· arXiv (Cornell University)· Computer Science
machine prediction:candidate · noneconsensus · none
3
citations
affunlabeled
Einstein's cosmological considerations
Daryl Janzen
2014· preprint· en· arXiv (Cornell University)· Physics and Astronomy
machine prediction:candidate · noneconsensus · none
3
citations

How this was built: Screen · Findings · About