Explaining regional variations in colon cancer survival in Ontario, Canada: a population-based retrospective cohort study
Bibliographic record
Abstract
OBJECTIVES: Regional variation in cancer survival is an important health system performance measurement. We evaluated if regional variation in colon cancer survival may be driven by differences in the patient population, their health and healthcare utilisation, and/or cancer care delivery. DESIGN: Population-based retrospective cohort study using routinely collected linked health administrative data. SETTING: Ontario, Canada. PARTICIPANTS: Patients with colon cancer diagnosed between 1 January 2009 and 31 December 2012. OUTCOME: Cancer-specific survival was compared across the province's 14 health regions. Using accelerated failure time models, we assessed whether regional survival variations were mediated through differences in case mix, including age, sex, comorbidities, stage at diagnosis and colon subsite, potential marginalisation and/or prediagnosis healthcare. RESULTS: The study population included 16 895 patients with colon cancer. There was statistically significant regional variation in cancer-specific survival. Three regions had cancer-specific survival that was between 30% (95% CI 1.03 to 1.65) and 39% (95% CI 1.13 to 1.71) longer and one region had cancer-specific survival that was 26% shorter (95% CI 0.58 to 0.93) than the reference region. For three of these regions, case mix explained between 26% and 56% of the survival variation. Further adjustment for rurality explained 22% of the remaining survival variation in one region. Adjustment for continuity of primary care and the diagnostic interval length explained 10% and 11% of the remaining survival variation in two other regions. Socioeconomic marginalisation, recent immigration and colonoscopy history did not explain colon cancer survival variation. CONCLUSIONS: Case mix accounted for much of the regional variation in colon cancer survival, indicating that efforts to monitor the quality of cancer care through survival metrics should consider case mix when reporting regional survival differences. Future work should repeat this approach in other settings and other cancer sites considering a broad range of potential mediators.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.004 |
| Meta-epidemiology (narrow) | 0.000 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.004 |
| Science and technology studies | 0.002 | 0.001 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".