Chromatin organisation and cancer prognosis: a pan-cancer study
Bibliographic record
Abstract
BACKGROUND: Chromatin organisation affects gene expression and regional mutation frequencies and contributes to carcinogenesis. Aberrant organisation of DNA has been correlated with cancer prognosis in analyses of the chromatin component of tumour cell nuclei using image texture analysis. As yet, the methodology has not been sufficiently validated to permit its clinical application. We aimed to define and validate a novel prognostic biomarker for the automatic detection of heterogeneous chromatin organisation. METHODS: Machine learning algorithms analysed the chromatin organisation in 461 000 images of tumour cell nuclei stained for DNA from 390 patients (discovery cohort) treated for stage I or II colorectal cancer at the Aker University Hospital (Oslo, Norway). The resulting marker of chromatin heterogeneity, termed Nucleotyping, was subsequently independently validated in six patient cohorts: 442 patients with stage I or II colorectal cancer in the Gloucester Colorectal Cancer Study (UK); 391 patients with stage II colorectal cancer in the QUASAR 2 trial; 246 patients with stage I ovarian carcinoma; 354 patients with uterine sarcoma; 307 patients with prostate carcinoma; and 791 patients with endometrial carcinoma. The primary outcome was cancer-specific survival. FINDINGS: In all patient cohorts, patients with chromatin heterogeneous tumours had worse cancer-specific survival than patients with chromatin homogeneous tumours (univariable analysis hazard ratio [HR] 1·7, 95% CI 1·2-2·5, in the discovery cohort; 1·8, 1·0-3·0, in the Gloucester validation cohort; 2·2, 1·1-4·5, in the QUASAR 2 validation cohort; 3·1, 1·9-5·0, in the ovarian carcinoma cohort; 2·5, 1·8-3·4, in the uterine sarcoma cohort; 2·3, 1·2-4·6, in the prostate carcinoma cohort; and 4·3, 2·8-6·8, in the endometrial carcinoma cohort). After adjusting for established prognostic patient characteristics in multivariable analyses, Nucleotyping was prognostic in all cohorts except for the prostate carcinoma cohort (HR 1·7, 95% CI 1·1-2·5, in the discovery cohort; 1·9, 1·1-3·2, in the Gloucester validation cohort; 2·6, 1·2-5·6, in the QUASAR 2 cohort; 1·8, 1·1-3·0, for ovarian carcinoma; 1·6, 1·0-2·4, for uterine sarcoma; 1·43, 0·68-2·99, for prostate carcinoma; and 1·9, 1·1-3·1, for endometrial carcinoma). Chromatin heterogeneity was a significant predictor of cancer-specific survival in microsatellite unstable (HR 2·9, 95% CI 1·0-8·4) and microsatellite stable (1·8, 1·2-2·7) stage II colorectal cancer, but microsatellite instability was not a significant predictor of outcome in chromatin homogeneous (1·3, 0·7-2·4) or chromatin heterogeneous (0·8, 0·3-2·0) stage II colorectal cancer. INTERPRETATION: The consistent prognostic prediction of Nucleotyping in different biological and technical circumstances suggests that the marker of chromatin heterogeneity can be reliably assessed in routine clinical practice and could be used to objectively assist decision making in a range of clinical settings. An immediate application would be to identify high-risk patients with stage II colorectal cancer who might have greater absolute benefit from adjuvant chemotherapy. Clinical trials are warranted to evaluate the survival benefit and cost-effectiveness of using Nucleotyping to guide treatment decisions in multiple clinical settings. FUNDING: The Research Council of Norway, the South-Eastern Norway Regional Health Authority, the National Institute for Health Research, and the Wellcome Trust.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".