MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Data Mining Algorithms and Applications
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

1,100 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
1,100 works in the cohort · of 4,299,418page 3 of 22

Labels cover 1 of 1,100 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 1,100 of 1,100 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

affno abstractunlabeled
FIsViz: A Frequent Itemset Visualizer
Carson K. Leung, Pourang Irani, Christopher L. Carmichael
2008· book-chapter· en· Lecture notes in computer science· Computer Science
machine prediction:candidate · noneconsensus · none
57
citations
affunlabeled
FOSHU
Philippe Fournier‐Viger, Souleymane Zida
2015· article· en· Computer Science
machine prediction:candidate · insufficient_payloadconsensus · none
56
citations
affno abstractunlabeled
Advanced Pattern Mining
Jiawei Han, Micheline Kamber, Jian Pei
2012· book-chapter· en· Elsevier eBooks· Computer Science
machine prediction:candidate · noneconsensus · none
56
citations
affunlabeled
Big Data Analysis and Mining
Carson K. Leung
2017· book-chapter· en· IGI Global eBooks· Computer Science
machine prediction:candidate · noneconsensus · none
54
citations
affno abstractunlabeled
Weighted frequent itemset mining over uncertain databases
Jerry Chun‐Wei Lin, Wensheng Gan, Philippe Fournier‐Viger, Tzung‐Pei Hong, Vincent S. Tseng
2015· article· en· Applied Intelligence· Computer Science
machine prediction:candidate · noneconsensus · none
54
citations
afffundunlabeled
A New Approach for Mining Correlated Frequent Subgraphs
Mohammad Ehsan Shahmi Chowdhury, Chowdhury Farhan Ahmed, Carson K. Leung
2021· article· en· ACM Transactions on Management Information Systems· Computer Science
machine prediction:candidate · noneconsensus · none
51
citations
affunlabeled
A fast Algorithm for mining fuzzy frequent itemsets
Jerry Chun‐Wei Lin, Ting Li, Philippe Fournier‐Viger, Tzung‐Pei Hong
2015· article· en· Journal of Intelligent & Fuzzy Systems· Computer Science
machine prediction:candidate · noneconsensus · none
49
citations
affunlabeled
Hierarchical Document Clustering
Benjamin C. M. Fung, Ke Wang, Martin Ester
2005· book-chapter· en· IGI Global eBooks· Computer Science
machine prediction:candidate · noneconsensus · none
45
citations
afffundno abstractunlabeled
$$L_1$$ L 1 splitting rules in survival forests
Hoora Moradian, Denis Larocque, François Bellavance
2016· article· en· Lifetime Data Analysis· Computer Science
machine prediction:candidate · noneconsensus · none
43
citations
affunlabeled
RHUPS
Yoonji Baek, Unil Yun, Heonho Kim, Hyoju Nam, Hyunsoo Kim, Jerry Chun‐Wei Lin +2 more
2021· article· en· ACM Transactions on Intelligent Systems and Technology· Computer Science
machine prediction:candidate · noneconsensus · none
43
citations
afffundunlabeled
Text Mining with n-gram Variables
Matthias Schonlau, Nick Guenther, Ilia Sucholutsky
2017· article· en· The Stata Journal Promoting communications on statistics and Stata· Computer Science
machine prediction:candidate · noneconsensus · none
40
citations
affno abstractunlabeled
Associative Classifiers for Medical Images
Maria-Luiza Antonie, Osmar R. Zai͏̈ane, Alexandru Coman
2003· book-chapter· en· Lecture notes in computer science· Computer Science
machine prediction:candidate · noneconsensus · none
39
citations
affunlabeled
An efficient algorithm for fuzzy frequent itemset mining
Tsu‐Yang Wu, Jerry Chun‐Wei Lin, Unil Yun, Chun-Hao Chen, Gautam Srivastava, Xianbiao Lv
2020· article· en· Journal of Intelligent & Fuzzy Systems· Computer Science
machine prediction:candidate · noneconsensus · none
39
citations

How this was built: Screen · Findings · About