MétaCan
Menu
Cohort builder

4,299,418 works, Canadian by any of four routes.

Every filter state is a URL; the URL is the query; the query is citable via /q/⟨hash⟩. The page, the API and the export parse the same parameters.

The current cohort, streamed from the database: every work column, the machine labels, the provisional scores, and the per-row validation status. Exports are capped at 100,000 rows. Mints a permanent /q/ link for this exact query. The same filters always produce the same link, whoever asks.

Search term
Author
Year range
Sort
Language
Type
Field
Venue
Topic
Reinforcement Learning in Robotics
Retraction
Abstract
Evidence source
Study design
Label agreement
Label status

Direct Codex and Gemma labels are unvalidated and sparse. Distilled predictions cover the full frame and are also unvalidated. Choose the evidence source explicitly; absence of a direct label is never a negative label.

affaffiliation
fundfunder
venuejournal
aboutaboutness

The four routes compose: require the funder route and exclude affiliation to get the funder-only stratum no affiliation-based frame ever sees.

1,145 results · 1 filter active ·
Results by year
20002025
Publication date
Categories
Machine labels · sparse coverage
Evidence
Language
Type
Citations
An unlabeled work is unknown, not a negative. Label coverage is reported on every query.
1,145 works in the cohort · of 4,299,418page 1 of 23

Labels cover 2 of 1,145 works in this cohort. The rest are unlabeled, which is not a negative label: the label table is sparse today and grows as labeling rounds land.

Distilled predictions cover 1,145 of 1,145 works in this cohort. Predictions are machine_predicted_unvalidated. The Gemma side is a direct model label for every work (title-only); the Codex side is a distilled, calibrated classifier. Candidate is the union; consensus is the intersection.

afffundunlabeled
Deep Reinforcement Learning That Matters
Peter Henderson, Riashat Islam, Philip Bachman, Joëlle Pineau, Doina Precup, David Meger
2018· article· en· Proceedings of the AAAI Conference on Artificial Intelligence· Computer Science
machine prediction:candidate · noneconsensus · none
1,504
citations
affunlabeled
An Introduction to Deep Reinforcement Learning
Vincent François-Lavet, Peter Henderson, Riashat Islam, Marc G. Bellemare, Joëlle Pineau
2018· article· en· Foundations and Trends® in Machine Learning· Computer Science
machine prediction:candidate · noneconsensus · none
1,251
citations
afffundunlabeled
The Option-Critic Architecture
Pierre‐Luc Bacon, Jean Harb, Doina Precup
2017· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
695
citations
affno abstractunlabeled
Natural actor–critic algorithms
Shalabh Bhatnagar, Richard Sutton, Mohammad Ghavamzadeh, Mark Lee
2009· article· en· Automatica· Computer Science
machine prediction:candidate · noneconsensus · none
569
citations
afffundunlabeled
Online Planning Algorithms for POMDPs
Stéphane Ross, Joëlle Pineau, S. Paquet, Brahim Chaib-draa
2008· article· en· Journal of Artificial Intelligence Research· Computer Science
machine prediction:candidate · noneconsensus · none
515
citations
affunlabeled
Deep learning, reinforcement learning, and world models
Yutaka Matsuo, Yann LeCun, Maneesh Sahani, Doina Precup, David Silver, Masashi Sugiyama +2 more
2022· review· en· Neural Networks· Computer Science
machine prediction:candidate · noneconsensus · none
493
citations
affno abstractunlabeled
A survey of point-based POMDP solvers
Guy Shani, Joëlle Pineau, Robert Kaplow
2012· article· en· Autonomous Agents and Multi-Agent Systems· Computer Science
machine prediction:candidate · noneconsensus · none
433
citations
affno abstractunlabeled
Learning Options in Reinforcement Learning
Martin Stolle, Doina Precup
2002· book-chapter· en· Lecture notes in computer science· Computer Science
machine prediction:candidate · noneconsensus · none
270
citations
affunlabeled
Model-Based Bayesian Exploration
Richard Dearden, Nir Friedman, David André
2013· article· en· arXiv (Cornell University)· Computer Science
machine prediction:candidate · noneconsensus · none
235
citations
affunlabeled
Bayesian Reinforcement Learning: A Survey
Mohammed Ghavamzadeh, Shie Mannor, Joëlle Pineau, Aviv Tamar
2015· article· en· Foundations and Trends® in Machine Learning· Computer Science
machine prediction:candidate · noneconsensus · none
223
citations
afffundunlabeled
The Option-Critic Architecture
Pierre‐Luc Bacon, Jean Harb, Doina Precup
2017· preprint· en· Proceedings of the AAAI Conference on Artificial Intelligence· Computer Science
machine prediction:candidate · noneconsensus · none
209
citations
affunlabeled
Opposition-Based Reinforcement Learning
Hamid R. Tizhoosh
2006· article· en· Journal of Advanced Computational Intelligence and Intelligent Informatics· Computer Science
machine prediction:candidate · noneconsensus · none
207
citations
affunlabeled
Imagination-Augmented Agents for Deep Reinforcement Learning
Sébastien Racanière, Théophane Weber, David Reichert, Lars Buesing, Arthur Guez, Danilo Jimenez Rezende +9 more
2017· article· en· arXiv (Cornell University)· Computer Science
machine prediction:candidate · noneconsensus · none
163
citations
affunlabeled
Bounded Finite State Controllers
Pascal Poupart, Craig Boutilier
2003· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
162
citations
affunlabeled
Incremental Natural Actor-Critic Algorithms
Shalabh Bhatnagar, Mohammad Ghavamzadeh, Mark Lee, Richard S. Sutton
2007· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
157
citations
affunlabeled
Fitted Q-iteration in continuous action-space MDPs
András Antos, Csaba Szepesvári, Rémi Munos
2007· article· en· HAL (Le Centre pour la Communication Scientifique Directe)· Computer Science
machine prediction:candidate · noneconsensus · none
149
citations
affunlabeled
Bayes-Adaptive POMDPs
Stéphane Ross, Brahim Chaib-draa, Joëlle Pineau
2007· article· en· Computer Science
machine prediction:candidate · noneconsensus · none
112
citations
affunlabeled
Regularized Policy Iteration
Amir massoud Farahmand, Mohammad Ghavamzadeh, Shie Mannor, Csaba Szepesvári
2008· article· en· PolyPublie (École Polytechnique de Montréal)· Computer Science
machine prediction:candidate · noneconsensus · none
108
citations
afffundunlabeled
Multi-Step Reinforcement Learning: A Unifying Algorithm
Kristopher De Asis, Juan Hernandez-Garcia, Gerhard Holland, Richard S. Sutton
2018· article· en· Proceedings of the AAAI Conference on Artificial Intelligence· Computer Science
machine prediction:candidate · noneconsensus · none
107
citations

How this was built: Screen · Findings · About