P069 Correlation Between Histologic Indices and Ulcerative Colitis Activity Measures Among Patients in the HICKORY (Etrolizumab) Open-Label Induction Cohort
Bibliographic record
Abstract
BACKGROUND: Cross-sectional studies in ulcerative colitis (UC) have shown at best a moderate association between histologic and clinical measures of disease activity, but few longitudinal studies have evaluated this relationship. As endoscopy and histology are both independent predictors of clinical outcomes in UC, there remains a need to assess these measures in parallel to demonstrate clinical benefit. HICKORY (NCT02100696) is an ongoing Phase 3 study evaluating etrolizumab in anti-tumor necrosis factor (aTNF)–experienced patients with moderate-to-severe UC. The correlation between histologic changes and established disease activity measures at end of induction (week 14) was assessed using data from the open-label induction (OLI) cohort of HICKORY. METHODS: Study analysis is based on data from patients in the OLI cohort who received ≥1 dose of etrolizumab 105 mg subcutaneously every 4 weeks during the 14-week induction phase. Baseline and week 14 biopsies were scored by 1 of 4 central readers using the Nancy histologic index (NHI) and the Robarts histopathology index (RHI) in patients who had active baseline histology (NHI >1 and RHI >3) and complete scoring at week 14 (n = 97). Binary week 14 histologic outcomes were characterized by presence or absence of neutrophils (NHI ≤1 or RHI ≤3 and Geboes subgrades 2B.0/3.0). Mayo Clinic score (MCS) endoscopic subscore (ES) was used to assess endoscopy. Pairwise associations were quantified by Spearman correlation (ρ; for correlation between change from baseline scores) and Cohen kappa coefficients (κ; for agreement among week 14 outcomes). ΔNHI and ΔRHI were compared to determine presence of a minimum clinically important difference (MCID) in MCS (∆MCS ≥3 from baseline). RESULTS: At week 14, 22% (21/97), 23% (22/97), and 8% (8/97) of patients achieved resolution of neutrophilic inflammation based on either NHI or RHI/Geboes, endoscopic improvement (ES ≤ 1), and endoscopic remission (ES = 0), respectively. Among patients with endoscopic improvement and endoscopic remission, neutrophilic resolution was achieved in 55% (12/22) and 75% (6/8) of patients, respectively. ΔNHI and ΔRHI were highly correlated (ρ = 0.91). There was weak to no association between ΔNHI/ΔRHI/ΔES and Δfecal calprotectin (ρ = –0.02 to 0.38), ΔC-reactive protein (ρ = 0.03 to 0.07), Δalbumin (ρ = –0.19 to –0.10), Δhemoglobin (ρ = –0.22 to –0.19), and Δsegmented neutrophils in the blood (ρ = –0.06 to 0.01). Weak correlations were observed between ΔNHI/ΔRHI and ΔES (ρ = 0.26–0.27), Δrectal bleeding (ρ = 0.24–0.28), and Δstool frequency (ρ = 0.40–0.42). Correlations between NHI, RHI/Geboes, and ES with symptomatic outcomes were weak (κ = 0.28–0.45). Difference in the mean grouped by achievement of ΔMCS ≥3 suggests MCIDs in ΔNHI and ΔRHI of 1.2 and 8.6, respectively. CONCLUSION(S): The analysis showed weak to moderate agreement between changes in histologic scores and changes in endoscopic scores, and weak to no agreement between changes in histologic scores and changes in laboratory results at week 14. There was a weak correlation between histologic scores and symptoms at the end of induction. MCID results suggest that both the NHI and RHI appear to effectively evaluate neutrophilic resolution, making the changes in score more clinically interpretable.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.003 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".