Dynamic evolution of translation initiation mechanisms in prokaryotes
Bibliographic record
Abstract
It is generally believed that prokaryotic translation is initiated by the interaction between the Shine-Dalgarno (SD) sequence in the 5' UTR of an mRNA and the anti-SD sequence in the 3' end of a 16S ribosomal RNA. However, there are two exceptional mechanisms, which do not require the SD sequence for translation initiation: one is mediated by a ribosomal protein S1 (RPS1) and the other used leaderless mRNA that lacks its 5' UTR. To understand the evolutionary changes of the mechanisms of translation initiation, we examined how universal the SD sequence is as an effective initiator for translation among prokaryotes. We identified the SD sequence from 277 species (249 eubacteria and 28 archaebacteria). We also devised an SD index that is a proportion of SD-containing genes in which the differences of GC contents are taken into account. We found that the SD indices varied among prokaryotic species, but were similar within each phylum. Although the anti-SD sequence is conserved among species, loss of the SD sequence seems to have occurred multiple times, independently, in different phyla. For those phyla, RPS1-mediated or leaderless mRNA-used mechanisms of translation initiation are considered to be working to a greater extent. Moreover, we also found that some species, such as Cyanobacteria, may acquire new mechanisms of translation initiation. Our findings indicate that, although translation initiation is indispensable for all protein-coding genes in the genome of every species, its mechanisms have dynamically changed during evolution.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".