Novel association approach for variable number tandem repeats (VNTRs) identifies DOCK5 as a susceptibility gene for severe obesity
Bibliographic record
Abstract
Variable number tandem repeats (VNTRs) constitute a relatively under-examined class of genomic variants in the context of complex disease because of their sequence complexity and the challenges in assaying them. Recent large-scale genome-wide copy number variant mapping and association efforts have highlighted the need for improved methodology for association studies using these complex polymorphisms. Here we describe the in-depth investigation of a complex region on chromosome 8p21.2 encompassing the dedicator of cytokinesis 5 (DOCK5) gene. The region includes two VNTRs of complex sequence composition which flank a common 3975 bp deletion, all three of which were genotyped by polymerase chain reaction and fragment analysis in a total of 2744 subjects. We have developed a novel VNTR association method named VNTRtest, suitable for association analysis of multi-allelic loci with binary and quantitative outcomes, and have used this approach to show significant association of the DOCK5 VNTRs with childhood and adult severe obesity (P(empirical)= 8.9 × 10(-8) and P= 3.1 × 10(-3), respectively) which we estimate explains ~0.8% of the phenotypic variance. We also identified an independent association between the 3975 base pair (bp) deletion and obesity, explaining a further 0.46% of the variance (P(combined)= 1.6 × 10(-3)). Evidence for association between DOCK5 transcript levels and the 3975 bp deletion (P= 0.027) and both VNTRs (P(empirical)= 0.015) was also identified in adipose tissue from a Swedish family sample, providing support for a functional effect of the DOCK5 deletion and VNTRs. These findings highlight the potential role of DOCK5 in human obesity and illustrate a novel approach for analysis of the contribution of VNTRs to disease susceptibility through association studies.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".