MétaCan
Menu
Retour à la cohorte
Enregistrement W2023159692 · doi:10.1097/brs.0b013e3181e41f87

The Highest Level of Evidence in a High Impact Journal: Is This the Final Verdict?

2010· article· en· W2023159692 sur OpenAlexaff
Charles G. Fisher, Alexander R. Vaccaro

Notice bibliographique

RevueSpine · 2010
Typearticle
Langueen
DomaineMedicine
ThématiqueSpinal Fractures and Fixation Techniques
Établissements canadiensUniversity of British Columbia
Organismes subventionnairesnon disponible
Mots-clésMedicineExternal validityInternal validityFace validityPopulationRandomized controlled trialClinical trialConfoundingEvidence-based medicinePhysical therapyAlternative medicineClinical psychologySurgerySocial psychologyPsychometricsPsychologyInternal medicinePathology

Résumé

récupéré en direct d'OpenAlex

Evidence-based medicine involves the deliberate integration of clinical research into therapeutic decision making.1 A prospective, randomized control trial (PRCT) is assumed to equally distribute unknown confounding variables and only manipulate the “treatment variable.” In medical PRCTs, this usually occurs in a well-defined population to determine efficacy; that is, does the intervention work on the participants in the study (internal validity)? In most medical (drug) studies, the greater challenge is determining whether the results can be extended to patients not in the study, so called external validity. Surgical or interventional studies face the same generalizabilty or external validity issues; however, one of the greatest challenges in surgical trials is patient recruitment, and the establishment of a valid study population to ensure internal validity. The recent history of medicine has been punctuated by PRCTs, which have established, reinforced, and challenged traditional clinical beliefs.2 One example of the astounding effect of level I trials is the recent literature on percutaneous vertebroplasty (PVP) compared with sham procedures3,4 and to conservative treatment5 for treatment of osteoporotic vertebral compression fractures. Many clinicians experience significant cognitive dissonance6 between the astounding early clinical improvement of many patients and the popular media perception of the results of these trials. An excellent review in this issue of the journal effectively dissects 2 of these studies and highlights strengths and weaknesses.5 However, the impact of these studies and others really distills down to the difficulties with establishing both internal and external validity. The major challenge of PRCTs is the requirement of equipoise by both patients and physicians regarding 2 apparently equally effective treatments. Clinicians may be biased to recommend direct interventions to some patients and only enroll patients in whom there is less severe intensity of symptoms, although this is controlled for once the study begins, the bias can be still appear through the inclusion-exclusion criteria developed by the clinicians. It is often the inclusion-exclusion measures that are adjusted based on expected recruitment. Patient consent to participate in interventional studies is by far the greatest challenge as systematic differences evolve between patients willing to participate in randomization. There may be significant differences in risk taking, expectations, and perseverance in people who are willing to relegate their treatment to chance versus patients who refused to participate. This volunteer bias would be akin to selection bias in observational studies. Most interventional studies have roughly 33% enrollment of those patients eligible. Patient preference is a big part of evidence-based medicine, and perhaps, the reason for the observational study advocates touting it as the best methodology for ensuring external validity. This dilemma was addressed in the Spine Patient Outcomes Research Trial by including an observational cohort to capture patients who were not enrolled in the prospective, randomized study.7,8 PRCTs are expected to have appropriate and accurate long term follow-up. However, in some situations long term follow-up may be less relevant than early functional results. For example, in the orthopedic literature, the natural history of most femur fractures is healing by 6 to 12 months regardless of treatment.9,10 The goal of internal fixation is early mobilization and pain control to avoid the sequelae of prolonged immobilization, possibly at the expense of soft tissue stripping and healing. Similarly, because the vertebral bodies have an excellent blood supply and soft tissue envelope, it is not surprising that the natural history of vertebral compression fractures is healing by 6 to 12 months. Is it fair to judge the long-term outcome of vertebroplasty compared with conservative treatment? Would anyone for go internal fixation of a femur fracture because of the equivocal long-term fracture healing? The real question is whether the “internal fixation” facilitates substantial early pain improvement and mobilization. Indeed, early pain relief after PVP has been observed in all of the PVP trials.11 An important overlooked result in the Rousing et al study was the significant (1.3 point, P < 0.02) improvement in Barthel functional score in the PVP group over the conservative group at 12 months. This functional improvement may be a reflection of earlier pain control and mobilization compared with conservatively treated patients. In practicing evidence-based medicine, we as clinicians and researchers must be wary of assigning truths to PRCTs, even if published in high impact journals. The so-called pinnacle of the research design hierarchy is not immune to methodological limitations often camouflaged by the unique and coveted PRCT design for evaluation of a surgical intervention. Common sense, continued investigation, and an appreciation of the principles around internal and external validity of high-level evidence studies will hopefully guide our practices in the future of spinal care.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction distillée sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.

score de la tête « metaresearch » (Codex)0,001
score de la tête « metaresearch » (Gemma)0,001
Version: codex-gemma-dda1882f352aStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Observationnel · Signal consensuel: aucune
GenreSignal candidat: Empirique · Signal consensuel: Empirique
Score de désaccord entre enseignants0,816
Score d'incertitude au seuil0,660

Scores Codex et Gemma par catégorie

CatégorieCodexGemma
Métarecherche0,0010,001
Méta-épidémiologie (sens strict)0,0000,000
Méta-épidémiologie (sens large)0,0000,000
Bibliométrie0,0000,000
Études des sciences et des technologies0,0000,000
Communication savante0,0000,000
Science ouverte0,0000,000
Intégrité de la recherche0,0000,001
Charge utile insuffisante (le modèle a refusé de juger)0,0010,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,202
Tête enseignante GPT0,418
Écart entre enseignants0,216 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeObservationnel
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations7
Publié2010
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueSpineMême sujetSpinal Fractures and Fixation TechniquesTravaux en français237 207