Evolution of Dinoflagellate Unigenic Minicircles and the Partially Concerted Divergence of Their Putative Replicon Origins
Bibliographic record
Abstract
Dinoflagellate chloroplast genes are unique in that each gene is on a separate minicircular chromosome. To understand the origin and evolution of this exceptional genomic organization we completely sequenced chloroplast psbA and 23S rRNA gene minicircles from four dinoflagellates: three closely related Heterocapsa species (H. pygmaea, H. rotundata, and H. niei) and the very distantly related Amphidinium carterae. We also completely sequenced a Protoceratium reticulatum minicircle with a 23S rRNA gene of novel structure. Comparison of these minicircles with those previously sequenced from H. triquetra and A. operculatum shows that in addition to the single gene all have noncoding regions of approximately a kilobase, which are likely to include a replication origin, promoter, and perhaps segregation sequences. The noncoding regions always have a high potential for folding into hairpins and loops. In all six dinoflagellate strains for which multiple minicircles are fully sequenced, parts of the noncoding regions, designated cores, are almost identical between the psbA and 23S rRNA minicircles, but the remainder is very different. There are two, three, or four cores per circle, sometimes highly related in sequence, but no sequence identity is detectable between cores of different species, even within one genus. This contrast between very high core conservation within a species, but none among species, indicates that cores are diverging relatively rapidly in a concerted manner. This is the first well-established case of concerted evolution of noncoding regions on numerous separate chromosomes. It differs from concerted evolution among tandemly repeated spacers between rRNA genes, and that of inverted repeats in plant chloroplast genomes, in involving only the noncoding DNA cores. We present two models for the origin of chloroplast gene minicircles in dinoflagellates from a typical ancestral multigenic chloroplast genome. Both involve substantial genomic reduction and gene transfer to the nucleus. One assumes differential gene deletion within a multicopy population of the resulting oligogenic circles. The other postulates active transposition of putative replicon origins and formation of minicircles by homologous recombination between them.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".