MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
scientometrics and bibliometrics research
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

2,100 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
2,100 works in the cohort · of 4,299,418page 1 of 42

Labels cover 193 of 2,100 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 2,100 of 2,100 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

affgemma · bibliometrics+metaresearchgpt · bibliometrics+scholarly_communicationmodels split
Bibliometrics: Methods for studying academic publishing
Anton Ninkov, Jason R. Frank, Lauren A. Maggio
2021· article· en· Perspectives on Medical Education· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · none
938
citations
affunlabeled
Citation Advantage of Open Access Articles
Günther Eysenbach
2006· article· en· PLoS Biology· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · none
830
citations
affno abstractunlabeled
Double-blind review favours increased representation of female authors
Amber E Budden, Tom Tregenza, Lonnie W. Aarssen, Julia Koricheva, Roosa Leimu, Christopher J. Lortie
2007· article· en· Trends in Ecology & Evolution· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · metaresearch
563
citations
afffundunlabeled
Team size matters: Collaboration and scientific impact since 1900
Vincent Larivière, Yves Gingras, Cassidy R. Sugimoto, Andrew Tsou
2014· article· en· Journal of the Association for Information Science and Technology· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · none
476
citations
afffundunlabeled
100 Most Cited Articles in Orthopaedic Surgery
Kelly A. Lefaivre, Babak Shadgan, Peter J. O’Brien
2010· article· en· Clinical Orthopaedics and Related Research· Decision Sciences
machine prediction:candidate · bibliometricsconsensus · none
365
citations
affunlabeled
Measuring the effectiveness of scientific gatekeeping
Kyle Siler, Kirby Lee, Lisa Bero
2014· article· en· Proceedings of the National Academy of Sciences· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · none
329
citations
affunlabeled
Contributorship and division of labor in knowledge production
Vincent Larivière, Nadine Desrochers, Benoît Macaluso, Philippe Mongeon, Adèle Paul‐Hus, Cassidy R. Sugimoto
2016· article· en· Social Studies of Science· Decision Sciences
machine prediction:candidate · metaresearch+bibliometrics+stsconsensus · none
264
citations
affunlabeled
Publish or impoverish
Bikun Chen, Fei Shu
2017· article· en· Aslib Journal of Information Management· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · none
255
citations
affgemma · metaresearch+scholarly_communicationgpt · metaresearch+research_integrity+scholarly_communicationmodels split
Firm action needed on predatory journals
Jocalyn Clark, Russell Smith
2015· editorial· en· BMJ· Decision Sciences
machine prediction:candidate · metaresearch+research_integrityconsensus · none
243
citations
affunlabeled
A simple proposal for the publication of journal citation distributions
Vincent Larivière, Véronique Kiermer, Catriona MacCallum, Marcia McNutt, Mark Patterson, Bernd Pulverer +3 more
2016· preprint· en· bioRxiv (Cold Spring Harbor Laboratory)· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · none
233
citations
affno abstractunlabeled
Measuring Technological Innovation over the Long Run
Bryan T. Kelly, Dimitris Papanikolaou, Amit Seru, Matt Taddy
2018· article· en· SSRN Electronic Journal· Decision Sciences
machine prediction:candidate · bibliometricsconsensus · none
230
citations
aboutno affunlabeled
Productivity, prominence, and the effects of academic environment
Samuel F. Way, Allison C. Morgan, Daniel B. Larremore, Aaron Clauset
2019· article· en· Proceedings of the National Academy of Sciences· Decision Sciences
machine prediction:candidate · metaresearch+bibliometricsconsensus · none
217
citations

How this was built: Screen · Findings · About