Progressive iron accumulation across multiple sclerosis phenotypes revealed by sparse classification of deep gray matter
Bibliographic record
Abstract
PURPOSE: To create an automated framework for localized analysis of deep gray matter (DGM) iron accumulation and demyelination using sparse classification by combining quantitative susceptibility (QS) and transverse relaxation rate (R2*) maps, for evaluation of DGM in multiple sclerosis (MS) phenotypes relative to healthy controls. MATERIALS AND METHODS: R2*/QS maps were computed using a 4.7T 10-echo gradient echo acquisition from 16 clinically isolated syndrome (CIS), 41 relapsing-remitting (RR), 40 secondary-progressive (SP), 13 primary-progressive (PP) MS patients, and 75 controls. Sparse classification for R2*/QS maps of segmented caudate nucleus (CN), putamen (PU), thalamus (TH), and globus pallidus (GP) structures produced localized maps of iron/myelin in MS patients relative to controls. Paired t-tests, with age as a covariate, were used to test for statistical significance (P ≤ 0.05). RESULTS: In addition to DGM structures found significantly different in patients compared to controls using whole region analysis, singular sparse analysis found significant results in RRMS PU R2* (P = 0.03), TH R2* (P = 0.04), CN QS (P = 0.04); in SPMS CN R2* (P = 0.04), GP R2* (P = 0.05); and in PPMS CN R2* (P = 0.04), TH QS (P = 0.04). All sparse regions were found to conform to an iron accumulation pattern of changes in R2*/QS, while none conformed to demyelination. Intersection of sparse R2*/QS regions also resulted in RRMS CN R2* becoming significant, while RRMS R2* TH and PPMS QS TH becoming insignificant. Common iron-associated volumes in MS patients and their effect size progressively increased with advanced phenotypes. CONCLUSION: A localized technique for identifying sparse regions indicative of iron or myelin in the DGM was developed. Progressive iron accumulation with advanced MS phenotypes was demonstrated, as indicated by iron-associated sparsity and effect size. LEVEL OF EVIDENCE: 1 Technical Efficacy: Stage 1 J. Magn. Reson. Imaging 2017;46:1464-1473.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".