Genetic and clinical correlates of two neuroanatomical AI dimensions in the Alzheimer’s disease continuum
Bibliographic record
Abstract
Alzheimer's disease (AD) is associated with heterogeneous atrophy patterns. We employed a semi-supervised representation learning technique known as Surreal-GAN, through which we identified two latent dimensional representations of brain atrophy in symptomatic mild cognitive impairment (MCI) and AD patients: the "diffuse-AD" (R1) dimension shows widespread brain atrophy, and the "MTL-AD" (R2) dimension displays focal medial temporal lobe (MTL) atrophy. Critically, only R2 was associated with widely known sporadic AD genetic risk factors (e.g., APOE ε4) in MCI and AD patients at baseline. We then independently detected the presence of the two dimensions in the early stages by deploying the trained model in the general population and two cognitively unimpaired cohorts of asymptomatic participants. In the general population, genome-wide association studies found 77 genes unrelated to APOE differentially associated with R1 and R2. Functional analyses revealed that these genes were overrepresented in differentially expressed gene sets in organs beyond the brain (R1 and R2), including the heart (R1) and the pituitary gland, muscle, and kidney (R2). These genes were enriched in biological pathways implicated in dendritic cells (R2), macrophage functions (R1), and cancer (R1 and R2). Several of them were "druggable genes" for cancer (R1), inflammation (R1), cardiovascular diseases (R1), and diseases of the nervous system (R2). The longitudinal progression showed that APOE ε4, amyloid, and tau were associated with R2 at early asymptomatic stages, but this longitudinal association occurs only at late symptomatic stages in R1. Our findings deepen our understanding of the multifaceted pathogenesis of AD beyond the brain. In early asymptomatic stages, the two dimensions are associated with diverse pathological mechanisms, including cardiovascular diseases, inflammation, and hormonal dysfunction-driven by genes different from APOE-which may collectively contribute to the early pathogenesis of AD. All results are publicly available at https://labs-laboratory.com/medicine/ .
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.003 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".