Changes in Crohnʼs disease phenotype over time in the Chinese population: Validation of the Montreal classification system
Bibliographic record
Abstract
BACKGROUND: Phenotypic evolution of Crohn's disease occurs in whites but has never been described in other populations. The Montreal classification may describe phenotypes more precisely. The aim of this study was to validate the Montreal classification through a longitudinal sensitivity analysis in detecting phenotypic variation compared to the Vienna classification. METHODS: This was a retrospective longitudinal study of consecutive Chinese Crohn's disease patients. All cases were classified by the Montreal classification and the Vienna classification for behavior and location. The evolution of these characteristics and the need for surgery were evaluated. RESULTS: A total of 109 patients were recruited (median follow-up: 4 years, range: 6 months-18 years). Crohn's disease behavior changed 3 years after diagnosis (P = 0.025), with an increase in stricturing and penetrating phenotypes, as determined by the Montreal classification, but was only detected by the Vienna classification after 5 years (P = 0.015). Disease location remained stable on follow-up in both classifications. Thirty-four patients (31%) underwent major surgery during the follow-up period with the stricturing [P = 0.002; hazard ratio (HR): 3.3; 95% CI: 1.5-7.0] and penetrating (P = 0.03; HR: 5.8; 95% CI: 1.2-28.2) phenotypes according to the Montreal classification associated with the need for major surgery. In contrast, colonic disease was protective against a major operation (P = 0.02; HR: 0.3; 95% CI: 0.08-0.8). CONCLUSIONS: This is the first study demonstrating phenotypic evolution of Crohn's disease in a nonwhite population. The Montreal classification is more sensitive to behavior phenotypic changes than is the Vienna classification after excluding perianal disease from the penetrating disease category and was useful in predicting course and the need for surgery.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".