MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Data Mining Algorithms and Applications
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

1,100 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
1,100 works in the cohort · of 4,299,418page 15 of 22

Labels cover 1 of 1,100 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 1,100 of 1,100 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

affno abstractunlabeled
Frequent and Non-frequent Sequential Itemsets Detection
Konstantinos F. Xylogiannopoulos, Panagiotis Karampelas, Reda Alhajj
2017· book-chapter· en· Lecture notes in social networks· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affno abstractunlabeled
Reasoning Engine for Support Maintenance
Rana Farah, Simon Hallé, Jiye Li, Freddy Lécué, B. Abeloos, Dominique Perron +5 more
2020· book-chapter· en· Lecture notes in computer science· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Reducing Subject Tree Browsing Complexity
Charles‐Antoine Julien, Pierre Tirilly, Jesse David Dinneen, Catherine Guastavino
2013· preprint· en· HAL (Le Centre pour la Communication Scientifique Directe)· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Repository Features to Help Researchers: An invitation to a dialogue
Chris Graf, Kiera McNeice, Wei Mun Chan, Sarah Callaghan, Ilaria Carnevale, Imogen Cranston +14 more
2021· article· en· Zenodo (CERN European Organization for Nuclear Research)· Computer Science
machine prediction:candidate · scholarly_communicationconsensus · none
1
citations
affno abstractunlabeled
Linear Algebra in Data Science
Peter Zizler, Roberta La Haye
2024· book· en· Compact textbooks in mathematics· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Data Mining with Incomplete Data
Hai Wang, Shouhong Wang
2008· book-chapter· en· IGI Global eBooks· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Pattern Discovery as Event Association
Andrew K. C. Wong, Yang Wang, Gary C.L. Li
2009· book-chapter· en· IGI Global eBooks· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affno abstractunlabeled
On Mining Maximal Pattern-Based Clusters
Jian Pei, Xiaoling Zhang, Moonjung Cho, Haixun Wang, Philip S. Yu
2008· book-chapter· en· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Introducción a Topic Modeling y MALLET
Shawn Graham, Scott Weingart, Ian Milligan
2018· article· es· The Programming Historian en español· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Finding Multidimensional Simpson's Paradox
Jay Xu, Jian Pei, Zicun Cong
2022· article· en· ACM SIGKDD Explorations Newsletter· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Association Bundle Identification
Wenxue Huang, Milorad Krneta, Li‐Min Lin
2009· book-chapter· en· IGI Global eBooks· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
fundno affno abstractunlabeled
Advances in Knowledge Discovery and Data Mining
De-Nian Yang, Xing Xie, Vincent S. Tseng, Jian Pei, Jen-Wei Huang, Jerry Chun‐Wei Lin
2024· book· en· Lecture notes in computer science· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Mining online shopping patterns and communities
Keivan Kianmehr, Xiao Peng, Chris Luce, Justin J. Chung, Nam Pham, W.K. Chung +3 more
2009· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affunlabeled
Three related types of multi-value association patterns
Thomas W.H. Lui, David Chiu
2008· article· en· Proceedings - International Conference on Pattern Recognition/Proceedings/International Conference on Pattern Recognition· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations
affno abstractunlabeled
Interactive Data Mining on a CBEA Cluster
Sabine McConnell, David R. Patton, Richard T. Hurley, Wilfred Blight, Graeme P. Young
2010· book-chapter· en· Lecture notes in computer science· Computer Science
machine prediction:candidate · noneconsensus · none
1
citations

How this was built: Screen · Findings · About