Revising the Suspected-Cancer Guidelines: Impacts on Patients’ Primary Care Contacts and Costs
Bibliographic record
Abstract
OBJECTIVES: This study aimed to explore the impact of revising suspected-cancer referral guidelines on primary care contacts and costs. METHODS: Participants had incident cancer (colorectal, n = 2000; ovary, n = 763; and pancreas, n = 597) codes in the Clinical Practice Research Datalink or England cancer registry. Difference-in-differences analyses explored guideline impacts on contact days and nonzero costs between the first cancer feature and diagnosis. Participants were controls ("old National Institute for Health and Care Excellence [NICE]") or "new NICE" if their index feature was introduced during guideline revision. Model assumptions were inspected visually and by falsification tests. Sensitivity analyses reclassified participants who subsequently presented with features in the original guidelines as "old NICE." For colorectal cancer, sensitivity analysis (n = 3481) adjusted for multimorbidity burden. RESULTS: Median contact days and costs were, respectively, 4 (interquartile range [IQR] 2-7) and £117.69 (IQR £53.23-£206.65) for colorectal, 5 (IQR 3-9) and £156.92 (IQR £78.46-£272.29) for ovary, and 7 (IQR 4-13) and £230.64 (IQR £120.78-£408.34) for pancreas. Revising ovary guidelines may have decreased contact days (incidence rate ratio [IRR] 0.74; 95% confidence interval 0.55-1.00; P = .05) with unchanged costs, but parallel trends assumptions were violated. Costs decreased by 13% (equivalent to -£28.05, -£50.43 to -£5.67) after colorectal guidance revision but only in sensitivity analyses adjusting for multimorbidity. Contact days and costs remained unchanged after pancreas guidance revision. CONCLUSIONS: The main analyses of symptomatic patients suggested that prediagnosis primary care costs remained unchanged after guidance revision for pancreatic cancer. For colorectal cancer, contact days and costs decreased in analyses adjusting for multimorbidity. Revising ovarian cancer guidelines may have decreased primary care contact days but not costs, suggesting increased resource-use intensity; nevertheless, there is evidence of confounding.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".