LINEs of evidence: noncanonical DNA replication as an epigenetic determinant
Bibliographic record
Abstract
LINE-1 (L1) retrotransposons are repetitive elements in mammalian genomes. They are capable of synthesizing DNA on their own RNA templates by harnessing reverse transcriptase (RT) that they encode. Abundantly expressed full-length L1s and their RT are found to globally influence gene expression profiles, differentiation state, and proliferation capacity of early embryos and many types of cancer, albeit by yet unknown mechanisms. They are essential for the progression of early development and the establishment of a cancer-related undifferentiated state. This raises important questions regarding the functional significance of L1 RT in these cell systems. Massive nuclear L1-linked reverse transcription has been shown to occur in mouse zygotes and two-cell embryos, and this phenomenon is purported to be DNA replication independent. This review argues against this claim with the goal of understanding the nature of this phenomenon and the role of L1 RT in early embryos and cancers. Available L1 data are revisited and integrated with relevant findings accumulated in the fields of replication timing, chromatin organization, and epigenetics, bringing together evidence that strongly supports two new concepts. First, noncanonical replication of a portion of genomic full-length L1s by means of L1 RNP-driven reverse transcription is proposed to co-exist with DNA polymerase-dependent replication of the rest of the genome during the same round of DNA replication in embryonic and cancer cell systems. Second, the role of this mechanism is thought to be epigenetic; it might promote transcriptional competence of neighboring genes linked to undifferentiated states through the prevention of tethering of involved L1s to the nuclear periphery. From the standpoint of these concepts, several hitherto inexplicable phenomena can be explained. Testing methods for the model are proposed.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.002 |
| Scholarly communication | 0.002 | 0.002 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.002 |
| Insufficient payload (model declined to judge) | 0.006 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".