Single-cell genomics unveils a canonical origin of the diverse mitochondrial genomes of euglenozoans
Bibliographic record
Abstract
BACKGROUND: The supergroup Euglenozoa unites heterotrophic flagellates from three major clades, kinetoplastids, diplonemids, and euglenids, each of which exhibits extremely divergent mitochondrial characteristics. Mitochondrial genomes (mtDNAs) of euglenids comprise multiple linear chromosomes carrying single genes, whereas mitochondrial chromosomes are circular non-catenated in diplonemids, but circular and catenated in kinetoplastids. In diplonemids and kinetoplastids, mitochondrial mRNAs require extensive and diverse editing and/or trans-splicing to produce mature transcripts. All known euglenozoan mtDNAs exhibit extremely short mitochondrial small (rns) and large (rnl) subunit rRNA genes, and absence of tRNA genes. How these features evolved from an ancestral bacteria-like circular mitochondrial genome remains unanswered. RESULTS: We sequenced and assembled 20 euglenozoan single-cell amplified genomes (SAGs). In our phylogenetic and phylogenomic analyses, three SAGs were placed within kinetoplastids, 14 within diplonemids, one (EU2) within euglenids, and two SAGs with nearly identical small subunit rRNA gene (18S) sequences (EU17/18) branched as either a basal lineage of euglenids, or as a sister to all euglenozoans. Near-complete mitochondrial genomes were identified in EU2 and EU17/18. Surprisingly, both EU2 and EU17/18 mitochondrial contigs contained multiple genes and one tRNA gene. Furthermore, EU17/18 mtDNA possessed several features unique among euglenozoans including full-length rns and rnl genes, six mitoribosomal genes, and nad11, all likely on a single chromosome. CONCLUSIONS: Our data strongly suggest that EU17/18 is an early-branching euglenozoan with numerous ancestral mitochondrial features. Collectively these data contribute to untangling the early evolution of euglenozoan mitochondria.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".