MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Software Engineering Research
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

3,468 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
3,468 works in the cohort · of 4,299,418page 43 of 70

Labels cover 10 of 3,468 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 3,468 of 3,468 works in this cohort. Predictions are machine_predicted_unvalidated teacher distillation outputs. Candidate is the union; consensus is the intersection.

affunlabeled
End-to-End Rationale Reconstruction
Mouna Dhaouadi, Bentley Oakes, Michalis Famelis
2022· preprint· en· Computer Science
distilled prediction:candidate · insufficient_payloadconsensus · none
5
citations
afffundunlabeled
UML Consistency Rules
Damiano Torre, Yvan Labiche, Marcela Genero, Maged Elaasar, Claudio Menghi
2020· article· en· Computer Science
distilled prediction:candidate · insufficient_payloadconsensus · none
5
citations
affunlabeled
Test Smell Detection Tools: A Systematic Mapping Study
Wajdi Aljedaani, Anthony Peruma, Ahmed Aljohani, Mazen Alotaibi, Mohamed Wiem Mkaouer, Ali Ouni +3 more
2021· dataset· en· Zenodo (CERN European Organization for Nuclear Research)· Computer Science
distilled prediction:candidate · metaresearch+metaepi_narrow+sts+scholarly_communication+insufficient_payloadconsensus · insufficient_payload
5
citations
affunlabeled
Code repurposing as an assessment tool
Joseph Sant
2015· article· en· International Conference on Software Engineering· Computer Science
distilled prediction:candidate · metaepi_narrowconsensus · none
5
citations
afffundunlabeled
Analyzing Structural Security Posture to Evaluate System Design Decisions
Joe Samuel, Jason Jaskolka, George Yee
2021· article· en· 2021 IEEE 21st International Conference on Software Quality, Reliability and Security (QRS)· Computer Science
distilled prediction:candidate · metaresearch+metaepi_narrow+scholarly_communicationconsensus · none
5
citations
afffundgemma · metaresearchgpt · no categorymodels split
On the relationship between use cases and test suites size
Mourad Badri, Linda Badri, William Flageol
2013· article· en· ACM SIGSOFT Software Engineering Notes· Computer Science
distilled prediction:candidate · metaresearch+metaepi_narrowconsensus · none
5
citations
affunlabeled
Mining temporal properties of data invariants
Caroline Lemieux
2015· article· en· International Conference on Software Engineering· Computer Science
distilled prediction:candidate · noneconsensus · none
5
citations

How this was built: Screen · Findings · About