Proteome wide association studies of LRRK2 variants identify novel causal and druggable proteins for Parkinson’s disease
Bibliographic record
Abstract
Common and rare variants in the LRRK2 locus are associated with Parkinson's disease (PD) risk, but the downstream effects of these variants on protein levels remain unknown. We performed comprehensive proteogenomic analyses using the largest aptamer-based CSF proteomics study to date (7006 aptamers (6138 unique proteins) in 3107 individuals). The dataset comprised six different and independent cohorts (five using the SomaScan7K (ADNI, DIAN, MAP, Barcelona-1 (Pau), and Fundació ACE (Ruiz)) and the PPMI cohort using the SomaScan5K panel). We identified eleven independent SNPs in the LRRK2 locus associated with the levels of 25 proteins as well as PD risk. Of these, only eleven proteins have been previously associated with PD risk (e.g., GRN or GPNMB). Proteome-wide association study (PWAS) analyses suggested that the levels of ten of those proteins were genetically correlated with PD risk, and seven were validated in the PPMI cohort. Mendelian randomization analyses identified GPNMB, LCT, and CD68 causal for PD and nominate one more (ITGB2). These 25 proteins were enriched for microglia-specific proteins and trafficking pathways (both lysosome and intracellular). This study not only demonstrates that protein phenome-wide association studies (PheWAS) and trans-protein quantitative trail loci (pQTL) analyses are powerful for identifying novel protein interactions in an unbiased manner, but also that LRRK2 is linked with the regulation of PD-associated proteins that are enriched in microglial cells and specific lysosomal pathways.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".