MétaCan
Menu
Retour à la cohorte
Enregistrement W2560536885 · doi:10.1007/s11999-016-5198-0

Editorial: CORR ® Will Change to Double-blind Peer Review—What Took Us So Long to Get There?

2016· editorial· en· W2560536885 sur OpenAlexaboutno aff
Seth S. Leopold

Notice bibliographique

RevueClinical Orthopaedics and Related Research · 2016
Typeeditorial
Langueen
DomainePharmacology, Toxicology and Pharmaceutics
ThématiquePharmaceutical industry and healthcare
Établissements canadiensnon disponible
Organismes subventionnairesHarvard UniversityMayo Clinic
Mots-clésBlindingMedicineAlternative medicineAnnalsPeer reviewMEDLINEFamily medicineMedical educationRandomized controlled trialSurgeryLawPathology

Résumé

récupéré en direct d'OpenAlex

A common misconception about peer review in biomedical sciences is that most journals practice “double-blind” review, wherein the authors of papers under consideration do not know the reviewers’ identities, and reviewers do not know the authors’ names or institutions. While this is the practice among most orthopaedic journals of which I am aware, it is not normative nor is it even common among the better medical journals of the world. For example, the Journal of the American Medical Association (JAMA) family of journals practices single-blind peer review. The AMA Manual of Style, which guides the philosophies and practices of those journals, cites the challenges of achieving successful blinding as well as the lack of evidence supporting clear benefits of double-blinding as the reasons behind their preference for single-blind review [7]. In addition to JAMA and its many relatives, New England Journal of Medicine practices single-blind peer review, as do Lancet, Annals of Surgery (as well as Annals of Medicine), and Canadian Medical Association Journal [3]. A few journals even practice “open” peer review [1, 2] (or have tried and abandoned it [10]). Open peer review allows authors to know reviewers’ identities, and vice versa, with the hopes that it might result in fewer inflammatory comments from reviewers, and that it might inject an additional measure of transparency into the process. After all, reviewers can have conflicts of interest, too. But the evidence supporting any of these approaches remains inferential and indirect. For obvious reasons, it is not easy to conduct true experimental studies on this topic, and the few experiments that have been done by and large were either underpowered [5] or have focused on whether blinding influences the quality of the review [6, 8] rather than its result. The scant evidence we have on the latter point suggests that blinding makes little difference in manuscript disposition [12]. Because the available evidence suggested that blinding did not seem to matter, Clinical Orthopaedics and Related Research® has long allowed authors the choice of single- or double-blind peer review for the work they send us. In recent years, about half of the authors who have sent papers here have selected single-blind review, and about half have opted for a double-blind process. Because of this, our reviewers are accustomed to seeing papers both ways, which is unusual among biomedical journals. We therefore felt CORR® was the perfect setting for an experimental study that might provide more-definitive evidence to guide the practices that we and other journals use. In particular, we wished to determine whether the knowledge of a prestigious author's identity or institution might increase the likelihood that reviewers would recommend the work for publication. Other studies on the topic did not have the advantage of a cooperating journal in which both approaches to peer review were part of the journal's normal workflow. We saw this opportunity as too important to pass up, and so CORR's Board of Trustees endorsed conducting an experiment on the topic here. In the randomized trial conducted at CORR, and published recently in JAMA [11], two versions of a fabricated manuscript were sent out to several hundred peer reviewers. The two versions were identical except that in one version the names of well-known authors from prestigious institutions were visible to reviewers, while in the other version reviewers were blinded to authors’ identities and universities. The influence of prestige increased the likelihood a reviewer would recommend publication of the paper by about 20%. This difference seems meaningful, though perhaps not overwhelming if one considers that typically three reviewers evaluate each paper. The influence of author prestige might therefore change the result of about one review in five. Human nature being what it is, one might reasonably expect that the reputation of an author or institution should exert some pull on the peer-review process. In fact, I was surprised the differences were not more pronounced. But they were large enough that to keep things as fair as possible, the Senior Editor panel at CORR has decided that we will employ double-blind peer review for all scientific manuscripts here. We felt it important to have good-quality evidence before making a fundamental change to our external-review process. We now have that evidence. We do not expect this policy to be a panacea. Experience—as well as evidence [5, 8, 12]—suggests that reviewers often can identify authors even when manuscripts are blinded. Such unintended unblinding is likely to become more common as an increasing number of orthopaedic projects are registered prospectively in clinical-trial databases such as www.clinicaltrials.gov. Such registration recently became a requirement for randomized trials in several important general-interest journals of our specialty, including CORR [9]. The controversies on this topic, the experiment's somewhat-troubling findings [11], and the fact that even a thoughtfully arrived-at policy is unlikely to eliminate fully even this one kind of bias (from among the numerous others that certainly remain) highlight how very complicated peer review is, and how difficult it is to do well. Winston Churchill offered this observation on the subject of democracy: “Many forms of Government have been tried, and will be tried in this world of sin and woe. No one pretends that democracy is perfect or all-wise. Indeed, it has been said that democracy is the worst form of Government except all those other forms that have been tried from time to time” [4]. The same might be said for peer review. Acknowledgments I would like to thank CORR's Board of Trustees and Lee Beadling BA, its Managing Director, for allowing the use of journal resources to answer this important question. In addition, I am grateful to Daniel J. Berry MD (of the Mayo Clinic) and James H. Herndon MD, MBA (of Harvard University) for lending their “identities” to the research project [11]. Finally, I appreciate the guidance from CORR's panel of Senior Editors on the topics of whether and how to convert the findings of the experiment into editorial policy here; these Editors are Matthew B. Dobbs MD, Mark C. Gebhardt MD, Terence J. Gioe MD, Paul A. Manner MD, Raphaël Porcher PhD, Clare M. Rimnac PhD, and Montri D. Wongworawat MD. Finally, I am grateful to Kanu Okike MD for his suggestions on this essay, as well as Dr. Okike, Mininder S. Kocher MD, MPH, Kevin T. Hug MD, and Brian Robertson for their partnership on the study that provided the evidence for this policy change.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,015
score de la tête « metaresearch » (Gemma)0,083
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesMétarecherche
Catégories consensuellesaucune
DomaineSignal candidat: Évaluation · Signal consensuel: aucune
Devis d'étudeSignal candidat: Sans objet · Signal consensuel: Sans objet
GenreSignal candidat: Éditorial · Signal consensuel: Éditorial
Score de désaccord entre enseignants0,985
Score d'incertitude au seuil0,305

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,0150,083
Méta-épidémiologie (sens strict)0,0040,002
Méta-épidémiologie (sens large)0,0040,004
Bibliométrie0,0030,002
Études des sciences et des technologies0,0040,005
Communication savante0,0130,007
Science ouverte0,0050,002
Intégrité de la recherche0,0180,015
Charge utile insuffisante (le modèle a refusé de juger)0,0910,098

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,634
Tête enseignante GPT0,665
Écart entre enseignants0,031 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Devis d'étudeSans objet
DomaineÉvaluation
GenreÉditorial

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations5
Publié2016
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueClinical Orthopaedics and Related ResearchMême sujetPharmaceutical industry and healthcareTravaux en français237 207