MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Software Engineering Research
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

3,468 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
3,468 works in the cohort · of 4,299,418page 8 of 70

Labels cover 10 of 3,468 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 3,468 of 3,468 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

afffundunlabeled
BinSequence
He Huang, Amr Youssef, Mourad Debbabi
2017· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
71
citations
affunlabeled
Clone region descriptors
Ekwa Duala-Ekoko, Martin P. Robillard
2010· article· en· ACM Transactions on Software Engineering and Methodology· Computer Science
machine prediction:candidate · noneconsensus · none
70
citations
afffundunlabeled
Detecting fragile comments
Inderjot Kaur Ratol, Martin P. Robillard
2017· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
67
citations
affno abstractunlabeled
Recommending reference API documentation
Martin P. Robillard, Yam Bahadur Chhetri
2014· article· en· Empirical Software Engineering· Computer Science
machine prediction:candidate · noneconsensus · none
67
citations
affno abstractunlabeled
Cohesive and Isolated Development with Branches
Earl T. Barr, Christian Bird, Peter C. Rigby, Abram Hindle, Daniel M. Germán, Prémkumar Dévanbu
2012· book-chapter· en· Lecture notes in computer science· Computer Science
machine prediction:candidate · noneconsensus · none
67
citations
affunlabeled
Code fragment summarization
Annie T. T. Ying, Martin P. Robillard
2013· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
67
citations
affunlabeled
Making sense of online code snippets
Siddharth Subramanian, Reid Holmes
2013· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
66
citations
affunlabeled
Plugging-in visualization
Rob Lintern, Jeff Michaud, Margaret‐Anne Storey, Xiaomin Wu
2003· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
66
citations
affunlabeled
A multidimensional empirical study on refactoring activity
Nikolaos Tsantalis, Victor Guana, Eleni Stroulia, Abram Hindle
2013· article· en· Conference of the Centre for Advanced Studies on Collaborative Research· Computer Science
machine prediction:candidate · noneconsensus · none
65
citations

How this was built: Screen · Findings · About