CAG repeat length in RAI1 is associated with age at onset variability in spinocerebellar ataxia type 2 (SCA2)
Bibliographic record
Abstract
Spinocerebellar ataxia type 2 (SCA2) is an autosomal dominant disorder caused by the expansion of a polymorphic (CAG)(n) tract, which is translated into an expanded polyglutamine tract in the ataxin-2 protein. Although repeat length and age at disease onset are inversely related, approximately 50% of the age at onset variance in SCA2 remains unexplained. Other familial factors have been proposed to account for at least part of this remaining variance in the polyglutamine dis-orders. The ability of polyglutamine tracts to interact with each other, as well as the presence of intra-nuclear inclusions in other polyglutamine disorders, led us to hypothesize that other CAG-containing proteins may interact with expanded ataxin-2 and affect the rate of protein accumulation, and thus influence age at onset. To test this hypothesis, we used step-wise multiple linear regression to examine 10 CAG-containing genes for possible influences on SCA2 age at onset. One locus, RAI1, contributed an additional 4.1% of the variance in SCA2 age at onset after accounting for the effect of the SCA2 expanded repeat. This locus was further studied in SCA3/Machado-Joseph disease (MJD), but did not have an effect on SCA3/MJD age at onset. This result implicates RAI1 as a possible contributor to SCA2 neurodegeneration and raises the possibility that other CAG-containing proteins may play a role in the pathogenesis of other polyglutamine disorders.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".