Quality of clinical practice guidelines in delirium: a systematic appraisal
Bibliographic record
Abstract
OBJECTIVE: To determine the accessibility and currency of delirium guidelines, guideline summary papers and evaluation studies, and critically appraise guideline quality. DESIGN: Systematic literature search for formal guidelines (in English or French) with focus on delirium assessment and/or management in adults (≥18 years), guideline summary papers and evaluation studies.Full appraisal of delirium guidelines published between 2008 and 2013 and obtaining a 'Rigour of Development' domain screening score cut-off of >40% using the Appraisal of Guidelines for Research and Evaluation (AGREE II) instrument. DATA SOURCES: Multiple bibliographic databases, guideline organisation databases, complemented by a grey literature search. RESULTS: 3327 database citations and 83 grey literature links were identified. A total of 118 retrieved delirium guidelines and related documents underwent full-text screening. A final 21 delirium guidelines (with 10 being >5 years old), 12 guideline summary papers and 3 evaluation studies were included. For 11 delirium guidelines published between 2008 and 2013, the screening AGREE II 'Rigour' scores ranged from 3% to 91%, with seven meeting the cut-off score of >40%. Overall, the highest rating AGREE II domains were 'Scope and Purpose' (mean 80.1%, range 64-100%) and 'Clarity and Presentation' (mean 76.7%, range 38-97%). The lowest rating domains were 'Applicability' (mean 48.7%, range 8-81%) and 'Editorial Independence' (mean 53%, range 2-90%). The three highest rating guidelines in the 'Applicability' domain incorporated monitoring criteria or audit and costing templates, and/or implementation strategies. CONCLUSIONS: Delirium guidelines are best sourced by a systematic grey literature search. Delirium guideline quality varied across all six AGREE II domains, demonstrating the importance of using a formal appraisal tool prior to guideline adaptation and implementation into clinical settings. Adding more knowledge translation resources to guidelines may improve their practical application and effective monitoring. More delirium guideline evaluation studies are needed to determine their effect on clinical practice.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.010 | 0.781 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".