MétaCan
Menu
Back to cohort
Record W7115728806 · doi:10.48448/t85w-fd88

Reminding Peer Reviewers to Comment on Reporting Items as Instructed by the Journal? An Analysis of 2 Randomized Trials

2025· other· W7115728806 on OpenAlexaboutno aff

Bibliographic record

VenueUnderline Science Inc. · 2025
Typeother
Language
Field
Topic
Canadian institutionsnot available
Fundersnot available
KeywordsRandomized controlled trialIntervention (counseling)Peer reviewMEDLINEPeer groupMeta-analysisPsychological intervention

Abstract

fetched live from OpenAlex

Hillary Wnfried Ramirez,<sup>1,2,3</sup> Malena Chiaborelli,<sup>1,2,3</sup> Christof M. Schönenberger,<sup>1</sup> Katie Mellor,<sup>4,5</sup> Alexandra N. Griessbach,<sup>1</sup> Paula Dhiman,<sup>4,6</sup> Pooja Gandhi,<sup>7</sup> Szimonetta Lohner,<sup>8,9</sup> Arnav Agarwal,<sup>10,11</sup> Ayodele Odutayo,<sup>5,12</sup> Michael M. Schlussel,<sup>4,6</sup> Philippe Ravaud,<sup>13,14</sup> David Moher,<sup>15,16 </sup>Matthias Briel,<sup>1,10</sup> Isabelle Boutron,<sup>13,14</sup> Sally Hopewell,<sup>4</sup> Sara Schroter,<sup>17,18 </sup>Benjamin Speich<sup>1,4</sup> <h4>Objective </h4> Two randomized controlled trials (RCTs) conducted at journal level have shown that reminding peer reviewers about the 10 most important and underreported reporting items did not improve the reporting quality in published articles.<sup>1</sup> With this pooled in-depth analysis of peer reviewer reports, we aimed to assess at what stage the intervention failed. <h4>Design</h4> A subsample of peer reviewer reports from the control group (receiving no reminder) and the intervention group (receiving a reminder of the 10 most important reporting items) were analyzed. In brief, 2 blinded authors independently extracted from peer reviewer reports how many of the 10 key reporting items were flagged by peer reviewers for clarification. The main outcome of this analysis was the mean proportion of the 10 selected reporting items for which at least 1 peer reviewer requested clarification, assessed at the manuscript level. Furthermore, we assessed how many requested changes were later adequately reported in published articles. <h4>Results </h4> Across the RCTs, we had access to peer reviewer reports for 533 manuscripts (265 in the intervention group, assessing comments from 740 peer reviewers; and 268 in the control group, assessing comments from 719 peer reviewers). Our results indicate that reviewers in the intervention group requested clarification on more reporting items than those in the control group. Overall, reviewers in the intervention group flagged 21.1% of the 10 reporting items for clarification compared with 13.1% in the control group (mean difference, 8.0 percentage points (pp); 95% CI, 4.9-11.1 pp). However, the overall mean difference between groups was diluted from 8.0 to 4.2 pp when only assessing accepted and published articles and decreased even further to 2.6 pp when only considering changes that were then implemented by authors of manuscripts (<b>Table 25-0956</b>). Approximately 55% of reporting items that were criticized by peer reviewers were later adequately reported in the published article (intervention group, 173 of 308 [56.2%]; control group, 105 of 197 [53.3%]). https://assets.underline.io/markdown_image/1/image/80fe50628cb162743e45faadd6604b8b.png <h4>Conclusions</h4> Reminding peer reviewers to check reporting items increased their focus on reporting guidelines, leading to more reporting-related requests in their review reports. However, the effect was diluted during the peer review process (particularly due to rejected articles and requests not being implemented by authors). Journals should therefore make sure that requested clarifications are adequately addressed in revised manuscripts. <h4>Reference</h4> 1. Speich B, Mann E, Schönenberger CM, et al. Reminding peer reviewers of reporting guideline items to improve completeness in published articles: primary results of 2 randomized trials. <i>JAMA Netw Open</i>. 2023;6(6):e2317651. <sup>1</sup>CLEAR Methods Center, Division of Clinical Epidemiology, Department Clinical Research, University Hospital Basel, University of Basel, Basel, Switzerland, benjamin.speich@usb.ch; <sup>2</sup>Swiss Tropical and Public Health Institute, Basel, Switzerland; <sup>3</sup>University of Basel, Basel, Switzerland; <sup>4</sup>Centre for Statistics in Medicine, Nuffield Department of Orthopaedics, Rheumatology and Musculoskeletal Sciences, University of Oxford, Oxford, UK; <sup>5</sup>Clinical Outcomes Assessment, Clarivate, London, UK; <sup>6</sup>The EQUATOR Network, Oxford, UK; <sup>7</sup>Department of Communication Sciences and Disorders, Faculty of Rehabilitation Medicine, University of Alberta, Edmonton, AL, Canada; <sup>8</sup>Cochrane Hungary, Medical School, University of Pécs, Pécs, Hungary; <sup>9</sup>MTA–PTE Lendület “Momentum” Evidence in Medicine Research Group, Department of Public Health Medicine, Medical School, University of Pécs, Pécs, Hungary; <sup>10</sup>Department of Health Research Methods, Evidence, and Impact, McMaster University, Hamilton, ON, Canada; <sup>11</sup>Division of General Internal Medicine, Department of Medicine, McMaster University, Hamilton, ON, Canada; <sup>12</sup>Division of Nephrology, Toronto General Hospital, University Health Network, Toronto, ON, Canada; <sup>13</sup>Centre d’Épidémiologie Clinique, Hôpital Hôtel-Dieu, Assistance Publique Hôpitaux de Paris, Paris, France;<sup>14</sup>Université de Paris, CRESS, Inserm, INRA, Paris, France; <sup>15</sup>Centre for Journalology, Clinical Epidemiology Program, Ottawa Hospital Research Institute, Ottawa, Ontario, Canada; <sup>16</sup>Faculty of Medicine, School of Epidemiology and Public Health, University of Ottawa, Ottawa, Ontario, Canada; <sup>17</sup><i>The BMJ</i>, London, UK; <sup>18</sup>Faculty of Public Health &amp; Policy, London School of Hygiene &amp; Tropical Medicine, London, UK. <h4>Conflict of Interest Disclosures</h4> Benjamin Speich and Matthias Briel reported receiving unrestricted grants from Moderna for studies unrelated to the presented work. Sara Schroter is employed by BMJ Publishing Group. Katie Mellor is employed by Clarivate. David Moher, Sally Hopewell, and Isabelle Boutron are members of the Consolidated Standards for Reporting Trials (CONSORT) executive board and authors of the CONSORT 2010 statement. David Moher is an author of the Standard Protocol Items: Recommendations for Interventional Trials (SPIRIT) 2013 statement. David Moher, Michael M. Schlussel, Paula Dhiman, and Philippe Ravaud are members of the Enhancing the Quality and Transparency of Research (EQUATOR) network. Isabelle Boutron and David Moher are members of the Peer Review Congress Advisory Board but were not involved in the review or decision for this abstract. No other disclosures were reported. <h4>Funding/Support </h4> Benjamin Speich was supported by a Return Postdoc.Mobility (P4P4PM194496) grant from the Swiss National Science Foundation. Christof M. Schönenberger was funded by the Janggen Pöhn Foundation and the Swiss National Science Foundation (MD-PhD grant No. 323530221860). Szimonetta Lohner was supported by the Hungarian Academy of Sciences (MTA) within the framework of the Lendület Programme. <h4>Role of the Funder/Sponsor</h4> The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript. <h4>Additional Information</h4> Both trials (CONSORT-PR and SPIRIT-PR) were prospectively registered on Open Science Framework (<a href="https://osf.io/c4hn8"><span class="Hyperlink CharOverride-6">https://osf.io/c4hn8</span></a> and <a href="https://osf.io/z2hm9"><span class="Hyperlink CharOverride-6">https://osf.io/z2hm9</span></a>).

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.400
metaresearch head score (Gemma)0.465
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch, Meta-epidemiology (narrow), Meta-epidemiology (broad), Bibliometrics, Science and technology studies, Scholarly communication, Open science, Insufficient payload (model declined to judge)
Consensus categoriesMetaresearch, Meta-epidemiology (narrow), Bibliometrics, Science and technology studies
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: Not applicable
GenreCandidate signal: Other · Consensus signal: none
Teacher disagreement score0.353
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.4000.465
Meta-epidemiology (narrow)0.0020.001
Meta-epidemiology (broad)0.0150.003
Bibliometrics0.0140.027
Science and technology studies0.0030.005
Scholarly communication0.0020.001
Open science0.0060.001
Research integrity0.0010.002
Insufficient payload (model declined to judge)0.0080.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.095
GPT teacher head0.427
Teacher spread0.332 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; both teacher heads agree on what is shown here.

Study designNot applicable
Domainnot available
GenreOther

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2025
Admission routes1
Has abstractyes

Explore more

Same venueUnderline Science Inc.French-language works237,207