Assessing the quality of seven clinical practice guidelines from four professional regulatory bodies in Quebec: What's the verdict?
Bibliographic record
Abstract
RATIONALE, AIMS, AND OBJECTIVES: Clinical practice guidelines (CPGs) have become a common feature in the health and social care fields, as they promote evidence-based practice and aim to improve quality of care and patient outcome. However, the benefits of the recommendations reported in CPGs are only as good as the quality of the CPGs themselves. Indeed, rigorous development and strategies for reporting are significant precursors to successful implementation of the recommendations that are proposed. Unfortunately, research has demonstrated that there is much variability in their level of quality. Furthermore, the quality of many CPGs has yet to be examined. The aim of the present study was to assess the quality of seven CPGs from four Quebec professional regulatory bodies pertaining to clinical evaluations in the fields of medicine, psychoeducation, psychotherapy, and social work. METHODS: The seven Quebec CPGs were assessed by four trained appraisers using the Appraisal of Guidelines for Research and Evaluation II guideline evaluation tool. RESULTS: Results suggest that while some quality criteria were met, most were not, denoting that these CPGs are of sub-optimal quality. CONCLUSION: Our findings highlight that there is still a lot to be done in order to improve the rigour and transparency with which scientific evidence is assessed and applied when developing CPGs. Impacts regarding the implementation of these CPGs are discussed in light of their use in clinical practice.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.177 | 0.818 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.008 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.004 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".