Effect of CAT or AGG Interruptions and CpG Methylation on Nucleosome Assembly upon Trinucleotide Repeats on Spinocerebellar Ataxia, Type 1 and Fragile X Syndrome*
Bibliographic record
Abstract
Nucleosome packaging regulates many aspects of DNA metabolism and is thought to mediate genetic instability and transcription of expanded trinucleotide repeats. Both instability and transcription are sensitive to repeat length, tract purity, and CpG methylation. CAT or AGG interruptions within the (CAG)n or (CGG)n tracts of spinocerebellar ataxia, type 1 or fragile X syndrome, respectively, confer increased genetic stability to the repeats. We report the formation of nucleosomes on sequences containing pure and interrupted (CAG)n and (CGG)n repeats having lengths above and below the genetic stability thresholds. Increased lengths of pure repeats led to increased and decreased propensities for nucleosome assembly on the (CAG)n and (CGG)n repeats, respectively. CpG methylation of the CGG repeat further reduced assembly. CAT interruptions in (CAG)n tracts decreased nucleosome assembly. In contrast, AGG interruptions in (CGG)n tracts did not affect assembly by hypoacetylated histones. The latter observation was unaltered by CpG methylation of the repeats. However, nucleosome assembly by hyperacetylated histones on interrupted CGG tracts was increased relative to pure tracts and this effect was abolished by CpG methylation. Thus, CAT or AGG interruptions can modulate the ability of (CAG)n and (CGG) tracts to assemble into chromatin and the effect of the AGG interruptions is dependent upon both the methylation status of the DNA and the acetylation status of the histones. Compared with the genetically unstable pure repeats, both interruptions permit a propensity of nucleosome assembly closer to that of random (genetically stable) sequences, suggesting an association of nucleosome assembly of trinucleotide repeats and genetic instability.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".