Activation‐induced cytidine deaminase can target multiple topologies of double‐stranded DNA in a transcription‐independent manner
Bibliographic record
Abstract
Activation-induced cytidine deaminase (AID) mutates immunoglobulin genes and acts genome-wide. AID targets robustly transcribed genes, and purified AID acts on single-stranded (ss) but not double-stranded (ds) DNA oligonucleotides. Thus, it is believed that transcription is the generator of ssDNA for AID. Previous cell-free studies examining the relationship between transcription and AID targeting have employed a bacterial colony count assay wherein AID reverts an antibiotic resistance stop codon in plasmid substrates, leading to colony formation. Here, we established a novel assay where kb-long dsDNA of varying topologies is incubated with AID, with or without transcription, followed by direct sequencing. This assay allows for an unselected and in-depth comparison of mutation frequency and pattern of AID targeting in the absence of transcription or across a range of transcription dynamics. We found that without transcription, AID targets breathing ssDNA in supercoiled and, to a lesser extent, in relaxed dsDNA. The most optimal transcription only modestly enhanced AID action on supercoiled dsDNA in a manner dependent on RNA polymerase speed. These data suggest that the correlation between transcription and AID targeting may reflect transcription leading to AID-accessible breathing ssDNA patches naturally occurring in de-chromatinized dsDNA, as much as being due to transcription directly generating ssDNA.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".