Exploring the role of low-frequency and rare exonic variants in alcohol and tobacco use
Bibliographic record
Abstract
Alcohol and tobacco use are heritable phenotypes. However, only a small number of common genetic variants have been identified, and common variants account for a modest proportion of the heritability. Therefore, this study aims to investigate the role of low-frequency and rare variants in alcohol and tobacco use. We meta-analyzed ExomeChip association results from eight discovery cohorts and included 12,466 subjects and 7432 smokers in the analysis of alcohol consumption and tobacco use, respectively. The ExomeChip interrogates low-frequency and rare exonic variants, and in addition a small pool of common variants. We investigated top variants in an independent sample in which ICD-9 diagnoses of “alcoholism” (N = 25,508) and “tobacco use disorder” (N = 27,068) had been assessed. In addition to the single variant analysis, we performed gene-based, polygenic risk score (PRS), and pathway analyses. The meta-analysis did not yield exome-wide significant results. When we jointly analyzed our top results with the independent sample, no low-frequency or rare variants reached significance for alcohol consumption or tobacco use. However, two common variants that were present on the ExomeChip, rs16969968 (p = 2.39 × 10−7) and rs8034191 (p = 6.31 × 10−7) located in CHRNA5 and AGPHD1 at 15q25.1, showed evidence for association with tobacco use. Low-frequency and rare exonic variants with large effects do not play a major role in alcohol and tobacco use, nor does the aggregate effect of ExomeChip variants. However, our results confirmed the role of the CHRNA5-CHRNA3-CHRNB4 cluster of nicotinic acetylcholine receptor subunit genes in tobacco use.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".