Predicting knee osteoarthritis progression using neural network with longitudinal MRI radiomics, and biochemical biomarkers: A modeling study
Bibliographic record
Abstract
BACKGROUND: Knee osteoarthritis (KOA) worsens both structurally and symptomatically, yet no model predicts KOA progression using Magnetic Resonance Image (MRI) radiomics and biomarkers. This study aimed to develop and test the longitudinal Load-Bearing Tissue Radiomic plus Biochemical biomarker and Clinical variable Model (LBTRBC-M) to predict KOA progression. METHODS AND FINDINGS: Data from the Foundation of the National Institutes of Health Osteoarthritis Biomarkers Consortium were used. We selected 594 participants with Kellgren-Lawrence grades 1-3 and complete biomarker data. The mean age was 61.6 ± 8.9 years, 58.8% were female, and the racial distribution was 79.3% White or White, 18.0% Black or African American, and 2.7% Asian or other non-White. A total of 1,753 knee MRIs were included across the study period, comprising 594 at baseline, 575 at 1-year follow-up, and 584 at 2-year follow-up. Outcomes included (1) both Joint Space Narrowing (JSN) and pain progression (n = 567), (2) only JSN progression (n = 303), (3) only pain progression (n = 295), and (4) non-progression (JSN or pain) (n = 588), corresponding to an approximate ratio of 2:1:1:2. JSN progression was defined as a minimum joint space width (JSW) loss of ≥0.7 mm, and pain progression as a sustained (≥2 time points) increase of ≥9 points on the Western Ontario and McMaster Universities Osteoarthritis Index (WOMAC) pain subscale (0-100 scale). Using the eXtreme Gradient BOOSTing (XGBOOST) algorithm, the model was developed in the total development cohort (n = 877) and tested in the total test cohort (n = 876). In the total test cohort, the Area Under the receiver operating characteristic Curve (AUC) of LBTRBC-M for predicting JSN and pain progression, JSN progression, pain progression, and non-progression were 0.880 (95% confidence interval (CI) [0.853, 0.903]), 0.913 (95% CI [0.881, 0.937]), 0.886 (95% CI [0.856, 0.910]), and 0.909 (95% CI [0.888, 0.926]), respectively. The overall accuracy of LBTRBC-M was 70.1%. With LBTRBC-M assistance, the prognostic accuracy of resident physicians (n = 7) improved from 44.7%-49.0% to 64.4%-66.5%. The main limitations include the use of a non-routine MRI sequence, the lack of external validation in independent cohorts, and limited incorporation of all knee joint structures in radiomic feature extraction. CONCLUSIONS: In this study, we observed that longitudinal MRI radiomic features of load-bearing knee joint tissues provide potentially informative markers for predicting knee osteoarthritis progression. These findings may help guide future efforts toward early risk stratification and personalized management of KOA.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".