MétaCan
Menu
Retour à la cohorte
Enregistrement W2270659621

Grading the Strength of a Body of Evidence When Assessing Health Care Interventions for the Effective Health Care Program of the Agency for Healthcare Research and Quality: An Update

2013· article· en· W2270659621 sur OpenAlexaff
Nancy D Berkman, Kathleen N Lohr, Mohammed Ansari, Marian McDonagh, Ethan M. Balk, Evelyn P Whitlock, James Reston, Eric B Bass, Mary Butler, Gerald Gartlehner, Lisa Hartling, Robert L Kane, Melissa L McPheeters, Laura C Morgan, Sally C. Morton, Meera Viswanathan, Priyanka Sista, Stephanie Chang

Notice bibliographique

RevueEurope PMC (PubMed Central) · 2013
Typearticle
Langueen
DomaineEconomics, Econometrics and Finance
ThématiqueHealth Systems, Economic Evaluations, Quality of Life
Établissements canadiensUniversity of Alberta
Organismes subventionnairesnon disponible
Mots-clésSystematic reviewPsychological interventionGrading (engineering)Health careMedicineAgency (philosophy)MEDLINEEvidence-based medicineMedical educationNursingAlternative medicinePolitical scienceEngineering
DOInon disponible

Résumé

récupéré en direct d'OpenAlex

Systematic reviews are essential tools for summarizing information to help users make well-informed decisions about health care options. The Evidence-based Practice Center (EPC) program, supported by the Agency for Healthcare Research and Quality (AHRQ), produces substantial numbers of such reviews, including those that explicitly compare two or more clinical interventions (sometimes termed comparative effectiveness reviews). These reports synthesize a body of literature; the ultimate goal is to help clinicians, policymakers, and patients make well-considered decisions about health care. The goal of strength of evidence assessments is to provide clearly explained, well-reasoned judgments about reviewers’ confidence in their systematic review conclusions so that decisionmakers can use them effectively.Beginning in 2007, AHRQ supported a cross-EPC set of work groups to develop guidance on major elements of designing, conducting, and reporting systematic reviews. Together the materials form the EPC Methods Guide for Effectiveness and Comparative Effectiveness Reviews; one chapter focused on grading the strength of evidence. This chapter updates the original EPC strength of evidence approach, presenting findings and recommendations of a work group with experience in applying previous guidance; it should be considered current guidance for EPCs. The guidance applies primarily to systematic reviews of drugs, devices, and other preventive and therapeutic interventions; it may apply to exposures (characteristics or risk factors that are determinants of health outcomes) and broader health services research questions. It does not address reviews of medical tests.EPC reports support the work of many decisionmakers, but EPCs do not themselves develop recommendations or practice guidelines. In particular, we limit our grading strength of evidence approach to individual outcomes. Unlike grading systems that were designed to be used more directly by specific decisionmakers,– we do not develop global summary judgments of the relative benefits and harms of treatment comparisons.We briefly explore the rationale for grading strength of evidence, define domains of concern, and describe our recommended grading system for systematic reviews. The aims of this guidance are twofold: (1) to foster appropriate consistency and transparency in the methods that different EPCs use to grade strength of evidence and (2) to facilitate users’ interpretations of those grades for guideline development or other decisionmaking tasks. Because this field is rapidly evolving, future revisions are anticipated; they will reflect our increasing understanding and experience with the methodology.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,262
score de la tête « metaresearch » (Gemma)0,600
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesMétarecherche
Catégories consensuellesMétarecherche
DomaineSignal candidat: Évaluation · Signal consensuel: aucune
Devis d'étudeSignal candidat: Théorique ou conceptuel · Signal consensuel: aucune
GenreSignal candidat: Méthodes · Signal consensuel: aucune
Score de désaccord entre enseignants0,738
Score d'incertitude au seuil0,911

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,2620,600
Méta-épidémiologie (sens strict)0,0040,006
Méta-épidémiologie (sens large)0,0130,013
Bibliométrie0,0670,047
Études des sciences et des technologies0,0030,007
Communication savante0,0180,021
Science ouverte0,0100,010
Intégrité de la recherche0,0100,013
Charge utile insuffisante (le modèle a refusé de juger)0,0060,005

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,606
Tête enseignante GPT0,533
Écart entre enseignants0,073 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; l’étiquette directe de Gemma et le classifieur distillé Codex s’accordent sur ce qui est montré ici.

Devis d'étudeThéorique ou conceptuel
DomaineÉvaluation
GenreMéthodes

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations115
Publié2013
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueEurope PMC (PubMed Central)Même sujetHealth Systems, Economic Evaluations, Quality of LifeTravaux en français237 207