MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Topic Modeling
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

2,769 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
2,769 works in the cohort · of 4,299,418page 9 of 56

Labels cover 6 of 2,769 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 2,769 of 2,769 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

affunlabeled
Products of Hidden Markov Models.
Andrew D. Brown, Geoffrey E. Hinton
2001· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
33
citations
affunlabeled
A Probabilistic Answer Type Model
Christopher Pinchak, Dekang Lin
2006· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
31
citations
affunlabeled
Exploiting query logs for cross-lingual query suggestions
Wei Gao, Cheng Niu, Jian‐Yun Nie, Ming Zhou, Kam‐Fai Wong, Hsiao-Wuen Hon
2010· article· en· ACM Transactions on Information Systems· Computer Science
machine prediction:candidate · noneconsensus · none
31
citations
afffundno abstractunlabeled
Topic and sentiment aware microblog summarization for twitter
Syed Muhammad Ali, Zeinab Noorian, Ebrahim Bagheri, Chen Ding, Feras Al‐Obeidat
2018· article· en· Journal of Intelligent Information Systems· Computer Science
machine prediction:candidate · noneconsensus · none
31
citations
affno abstractunlabeled
Development of a generalizable natural language processing pipeline to extract physician-reported pain from clinical reports: Generated using publicly-available datasets and tested on institutional clinical reports for cancer patients with bone metastases
Hossein Naseri, Kamran Kafi, Sonia Skamene, Marwan Tolba, Mame Daro Faye, Paul Ramia +2 more
2021· article· en· Journal of Biomedical Informatics· Computer Science
machine prediction:candidate · noneconsensus · none
30
citations
affunlabeled
Protein Language Models: Is Scaling Necessary?
Quentin Fournier, Robert M. Vernon, Almer M. van der Sloot, B.M. Schulz, Sarath Chandar, Christopher J. Langmead
2024· preprint· en· bioRxiv (Cold Spring Harbor Laboratory)· Computer Science
machine prediction:candidate · noneconsensus · none
30
citations
fundno affunlabeled
Lightweight transformers for clinical natural language processing
Omid Rohanian, Mohammadmahdi Nouriborji, Hannah Jauncey, Samaneh Kouchaki, Farhad Nooralahzadeh, Lei Clifton +2 more
2024· article· en· Natural Language Engineering· Computer Science
machine prediction:candidate · noneconsensus · none
30
citations
affunlabeled
Extrinsic summarization evaluation
Gabriel Murray, Thomas Kleinbauer, Peter Poller, Tilman Becker, Steve Renals, Jonathan Kilgour
2009· article· en· ACM Transactions on Speech and Language Processing· Computer Science
machine prediction:candidate · noneconsensus · none
30
citations
affno abstractunlabeled
Towards Automatic Topical Question Generation
Yllias Chali, Sadid A. Hasan
2012· article· en· International Conference on Computational Linguistics· Computer Science
machine prediction:candidate · noneconsensus · none
29
citations
affunlabeled
Learning to Transfer Prompts for Text Generation
Junyi Li, Tianyi Tang, Jian‐Yun Nie, Ji-Rong Wen
2022· article· en· Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies· Computer Science
machine prediction:candidate · noneconsensus · none
28
citations

How this was built: Screen · Findings · About