Examining the environmental risk factors of progressive-onset and relapsing-onset multiple sclerosis: recruitment challenges, potential bias, and statistical strategies
Bibliographic record
Abstract
It is unknown whether the currently known risk factors of multiple sclerosis reflect the etiology of progressive-onset multiple sclerosis (POMS) as observational studies rarely included analysis by type of onset. We designed a case-control study to examine associations between environmental factors and POMS and compared effect sizes to relapse-onset MS (ROMS), which will offer insights into the etiology of POMS and potentially contribute to prevention and intervention practice. This study utilizes data from the Primary Progressive Multiple Sclerosis (PPMS) Study and the Australian Multi-center Study of Environment and Immune Function (the AusImmune Study). This report outlines the conduct of the PPMS Study, whether the POMS sample is representative, and the planned analysis methods. The study includes 155 POMS, 204 ROMS, and 558 controls. The distributions of the POMS were largely similar to Australian POMS patients in the MSBase Study, with 54.8% female, 85.8% POMS born before 1970, mean age of onset of 41.44 ± 8.38 years old, and 67.1% living between 28.9 and 39.4° S. The POMS were representative of the Australian POMS population. There are some differences between POMS and ROMS/controls (mean age at interview: POMS 55 years vs. controls 40 years; sex: POMS 53% female vs. controls 78% female; location of residence: 14.3% of POMS at a latitude ≤ 28.9°S vs. 32.8% in controls), which will be taken into account in the analysis. We discuss the methodological issues considered in the study design, including prevalence-incidence bias, cohort effects, interview bias and recall bias, and present strategies to account for it. Associations between exposures of interest and POMS/ROMS will be presented in subsequent publications.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".