Insights into the Evolutionary Origin and Genome Architecture of the Unicellular Opisthokonts <i>Capsaspora owczarzaki</i> and <i>Sphaeroforma arctica</i>
Bibliographic record
Abstract
Molecular phylogenetic analyses have recently shown that the unicellular amoeboid protist Capsaspora owczarzaki is unlikely to be a nucleariid or an ichthyosporean as previously described, but is more closely related to Metazoa, Choanoflagellata, and Ichthyosporea. However, the specific phylogenetic relationship of Capsaspora to other protist opisthokont lineages was poorly resolved. To test these earlier results we have expanded both the taxonomic sampling and the number of genes from opisthokont unicellular lineages. We have sequenced the protein-coding genes elongation factor 1-alpha (EF1-alpha) and heat shock protein 70 (Hsp70) from C. owczarzaki and the ichthyosporean Sphaeroforma arctica. Our maximum likelihood (ML) and Bayesian analyses of a concatenated alignment of EF1-alpha, Hsp70, and actin protein sequences with a better sampling of opisthokont-related protist lineages indicate that C. owczarzaki is not clearly allied with any of the unicellular opisthokonts, but represents an independent unicellular lineage closely related to animals and choanoflagellates. Moreover, we have found that the ichthyosporean S. arctica possesses an EF-like (EFL) gene copy instead of the canonical EF1-alpha, the first so far described in an ichthyosporean. A maximum likelihood phylogenetic analysis shows that the EF-like gene of S. arctica strongly groups with the EF-like genes from choanoflagellates. Finally, to begin characterizing the Capsaspora genome, we have performed pulsed-field gel electrophoresis (PFGE) analyses, which indicate that its genome has at least 12 chromosomes with a total genome size in the range of 22-25 Mb.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".