Copy Number Variation and Haplotype Analysis of <scp>17q21.31</scp> Reveals Increased Risk Associated with Progressive Supranuclear Palsy and Gene Expression Changes in Neuronal Cells
Bibliographic record
Abstract
BACKGROUND: The 17q21.31 region with various structural forms characterized by the H1/H2 haplotypes and three large copy number variations (CNVs) represents the strongest risk locus in progressive supranuclear palsy (PSP). OBJECTIVE: To investigate the association between CNVs and structural forms on 17q.21.31 with the risk of PSP. METHODS: Utilizing whole genome sequencing data from 1684 PSP cases and 2392 controls, the three large CNVs (α, β, and γ) and structural forms within 17q21.31 were identified and analyzed for their association with PSP. RESULTS: We found that the copy number of γ was associated with increased PSP risk (odds ratio [OR] = 1.10, P = 0.0018). From H1β1γ1 (OR = 1.21) and H1β2γ1 (OR = 1.24) to H1β1γ4 (OR = 1.57), structural forms of H1 with additional copies of γ displayed a higher risk for PSP. The frequency of the risk sub-haplotype H1c rises from 1% in individuals with two γ copies to 88% in those with eight copies. Additionally, γ duplication up-regulates expression of ARL17B, LRRC37A/LRRC37A2, and NSFP1, while down-regulating KANSL1. Single-nucleus RNA-seq of the dorsolateral prefrontal cortex analysis reveals γ duplication primarily up-regulates LRRC37A/LRRC37A2 in neuronal cells. CONCLUSIONS: The copy number of γ is associated with the risk of PSP after adjusting for H1/H2, indicating that the complex structure at 17q21.31 is an important consideration when evaluating the genetic risk of PSP. © 2025 The Author(s). Movement Disorders published by Wiley Periodicals LLC on behalf of International Parkinson and Movement Disorder Society.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".