Proteome wide association studies of LRRK2 variants identify novel causal and druggable proteins for Parkinson’s disease
Bibliographic record
Abstract
Common and rare variants in the LRRK2 locus are associated with Parkinson's disease (PD) risk, but the downstream effects of these variants on protein levels remain unknown. We performed comprehensive proteogenomic analyses using the largest aptamer-based CSF proteomics study to date (7006 aptamers (6138 unique proteins) in 3107 individuals). The dataset comprised six different and independent cohorts (five using the SomaScan7K (ADNI, DIAN, MAP, Barcelona-1 (Pau), and Fundació ACE (Ruiz)) and the PPMI cohort using the SomaScan5K panel). We identified eleven independent SNPs in the LRRK2 locus associated with the levels of 25 proteins as well as PD risk. Of these, only eleven proteins have been previously associated with PD risk (e.g., GRN or GPNMB). Proteome-wide association study (PWAS) analyses suggested that the levels of ten of those proteins were genetically correlated with PD risk, and seven were validated in the PPMI cohort. Mendelian randomization analyses identified GPNMB, LCT, and CD68 causal for PD and nominate one more (ITGB2). These 25 proteins were enriched for microglia-specific proteins and trafficking pathways (both lysosome and intracellular). This study not only demonstrates that protein phenome-wide association studies (PheWAS) and trans-protein quantitative trail loci (pQTL) analyses are powerful for identifying novel protein interactions in an unbiased manner, but also that LRRK2 is linked with the regulation of PD-associated proteins that are enriched in microglial cells and specific lysosomal pathways.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.003 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".