Detection of Slipped-DNAs at the Trinucleotide Repeats of the Myotonic Dystrophy Type I Disease Locus in Patient Tissues
Bibliographic record
Abstract
Slipped-strand DNAs, formed by out-of-register mispairing of repeat units on complementary strands, were proposed over 55 years ago as transient intermediates in repeat length mutations, hypothesized to cause at least 40 neurodegenerative diseases. While slipped-DNAs have been characterized in vitro, evidence of slipped-DNAs at an endogenous locus in biologically relevant tissues, where instability varies widely, is lacking. Here, using an anti-DNA junction antibody and immunoprecipitation, we identify slipped-DNAs at the unstable trinucleotide repeats (CTG)n•(CAG)n of the myotonic dystrophy disease locus in patient brain, heart, muscle and other tissues, where the largest expansions arise in non-mitotic tissues such as cortex and heart, and are smallest in the cerebellum. Slipped-DNAs are shown to be present on the expanded allele and in chromatinized DNA. Slipped-DNAs are present as clusters of slip-outs along a DNA, with each slip-out having 1-100 extrahelical repeats. The allelic levels of slipped-DNA containing molecules were significantly greater in the heart over the cerebellum (relative to genomic equivalents of pre-IP input DNA) of a DM1 individual; an enrichment consistent with increased allelic levels of slipped-DNA structures in tissues having greater levels of CTG instability. Surprisingly, this supports the formation of slipped-DNAs as persistent mutation products of repeat instability, and not merely as transient mutagenic intermediates. These findings further our understanding of the processes of mutation and genetic variation.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".