Copy Number Variant Risk Scores Associated With Cognition, Psychopathology, and Brain Structure in Youths in the Philadelphia Neurodevelopmental Cohort
Bibliographic record
Abstract
Importance: Psychiatric and cognitive phenotypes have been associated with a range of specific, rare copy number variants (CNVs). Moreover, IQ is strongly associated with CNV risk scores that model the predicted risk of CNVs across the genome. But the utility of CNV risk scores for psychiatric phenotypes has been sparsely examined. Objective: To determine how CNV risk scores, common genetic variation indexed by polygenic scores (PGSs), and environmental factors combine to associate with cognition and psychopathology in a community sample. Design, Setting, and Participants: The Philadelphia Neurodevelopmental Cohort is a community-based study examining genetics, psychopathology, neurocognition, and neuroimaging. Participants were recruited through the Children's Hospital of Philadelphia pediatric network. Participants with stable health and fluency in English underwent genotypic and phenotypic characterization from November 5, 2009, through December 30, 2011. Data were analyzed from January 1 through July 30, 2021. Exposures: The study examined (1) CNV risk scores derived from models of burden, predicted intolerance, and gene dosage sensitivity; (2) PGSs from genomewide association studies related to developmental outcomes; and (3) environmental factors, including trauma exposure and neighborhood socioeconomic status. Main Outcomes and Measures: The study examined (1) neurocognition, with the Penn Computerized Neurocognitive Battery; (2) psychopathology, with structured interviews based on the Schedule for Affective Disorders and Schizophrenia for School-Age Children; and (3) brain volume, with magnetic resonance imaging. Results: Participants included 9498 youths aged 8 to 21 years; 4906 (51.7%) were female, and the mean (SD) age was 14.2 (3.7) years. After quality control, 18 185 total CNVs greater than 50 kilobases (10 517 deletions and 7668 duplications) were identified in 7101 unrelated participants genotyped on Illumina arrays. In these participants, elevated CNV risk scores were associated with lower overall accuracy on cognitive tests (standardized β = 0.12; 95% CI, 0.10-0.14; P = 7.41 × 10-26); lower accuracy across a range of cognitive subdomains; increased overall psychopathology; increased psychosis-spectrum symptoms; and higher deviation from a normative developmental model of brain volume. Statistical models of developmental outcomes were significantly improved when CNV risk scores were combined with PGSs and environmental factors. Conclusions and Relevance: In this study, elevated CNV risk scores were associated with lower cognitive ability, higher psychopathology including psychosis-spectrum symptoms, and greater deviations from normative magnetic resonance imaging models of brain development. Together, these results represent a step toward synthesizing rare genetic, common genetic, and environmental factors to understand clinically relevant outcomes in youth.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".