P773 Patient reported outcomes, partial MAYO score and SCCAI are equally accurate in predicting mucosal healing in UC: Preliminary results from a prospective study
Bibliographic record
Abstract
Abstract Background Optimal management of patients with ulcerative colitis (UC) requires the accurate assessment of disease activity. Endoscopic evaluation is considered the gold standard approach, but it is invasive. We aimed to determine how strong patient reported outcomes, clinical scores and symptoms correlate with endoscopy for assessment of disease activity in UC patients. Methods One hundred and thirty-six patients were included prospectively (age: 48 (IQR: 38–61) years, duration 12 (4–19) years, 63 females, 53.7% extensive disease, 40.4% on biologicals) at the time of the colonoscopy. The 2 item patient reported outcome (PRO), partial MAYO, Simple Clinical Colitis Activity Index (SCCAI), Mayo endoscopic subscore (MES), Baron and Ulcerative Colitis Endoscopic Index of Severity (UCEIS) scores were calculated. C reactive Protein (CRP) and fecal calprotectin (FCAL) was available in 58.1 and 33.8% of patients. 20.7% had clinical flare, treatment was escalated in 17.8% of patients. Sensitivity, specificity, PPV and NPV values were calculated, ROC analysis and K-statistics were performed. Results Rectal bleeding(RBS), stool frequency(SF) subscore of 0, or total PRO2 remission(RBS0 and SF≤1), partial MAYO(≤2) and SCCAI(≤2.5) remission were similarly associated to mucosal healing defined by MES(0 or ≤1) or Baron (0 or ≤1) scores (Table 1). PRO2 remission (AUCMES0/Baron0:0.747/0.715, AUCMES0-1/Baron0-1:0.867/0.863), SF AUCMES0/Baron0:0.731/0.703, AUCMES0-1/Baron0-1:0851/0.839), RBS(AUCMES0/Baron0:0.708/0.685, AUCMES0-1/Baron0-1:0.828/0.835) partial Mayo (AUCMES0/Baron0:0.792/0.755, AUCMES0-1/Baron0-1:0.917/0.903) and SCCAI (AUCMES0/Baron0:0.738/0.724, AUCMES0-1/Baron0-1:0.908/0.880) were similarly associated with mucosal healing in a ROC analysis. There was a string association between MES and Baron (k=0.798), while moderate agreement between UCEIS and MES (K=0.451) or Baron (K=0.499) scores. Agreement between CRP and clinical remission or endoscopic healing (MES/Baron) was poor (K~0.2), while agreement between FCAL (>100 or >250) and RBS-PRO2 remission (K>250:0.56–0.61) or MES/Baron 0 was moderate to good(K>100:0.54-0.53 and K>250:0.50–0.54) Conclusion We found no difference across accuracy of RBS, SF, PRO2, partial Mayo and SCCAI in predicting endoscopic healing. A strong association was found with high PPV for MES/Baron ≤1 and high NPV for MES/Baron 0. FCAL, but not CRP was associated to clinical and endoscopic remission.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.006 | 0.011 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".