Bringing two worlds closer together: a critical analysis of an integrated approach to guideline development and quality assurance schemes
Bibliographic record
Abstract
BACKGROUND: Although quality indicators are frequently derived from guidelines, there is a substantial gap in collaboration between the corresponding parties. To optimise workflow, guideline recommendations and quality assurance should be aligned methodologically and practically. Learning from the European Commission Initiative on Breast Cancer (ECIBC), our objective was to bring the key knowledge and most important considerations from both worlds together to inform European Commission future initiatives. METHODS: We undertook several steps to address the problem. First, we conducted a feasibility study that included a survey, interviews and a review of manuals for an integrated guideline and quality assurance (QA) scheme that would support the European Commission. The feasibility study drew from an assessment of the ECIBC experience that followed commonly applied strategies leading to separation of the guideline and QA development processes. Secondly, we used results of a systematic review to inform our understanding of methodologies for integrating guideline and QA development. We then, in a third step, used the findings to prepare an evidence brief and identify key aspects of a methodological framework for integrating guidelines QA through meetings with key informants. RESULTS: Seven key themes emerged to be taken into account for integrating guidelines and QA schemes: (1) evidence-based integrated guideline and QA frameworks are possible, (2) transparency is key in clearly documenting the source and rationale for quality indicators, (3) intellectual and financial interests should be declared and managed appropriately, (4) selection processes and criteria for quality indicators need further refinement, (5) clear guidance on retirement of quality indicators should be included, (6) risks of an integrated guideline and QA Group can be mitigated, and (7) an extension of the GIN-McMaster Guideline Development Checklist should incorporate QA considerations. DISCUSSION: We concluded that the work of guideline and QA developers can be integrated under a common methodological framework and we provided key findings and recommendations. These two worlds, that are fundamental to improving health, can both benefit from integration.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.027 | 0.005 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.004 | 0.000 |
| Bibliometrics | 0.002 | 0.006 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".