Proposal of the reverse flow model for the origin of the eukaryotic cell based on comparative analyses of Asgard archaeal metabolism
Bibliographic record
Abstract
The origin of eukaryotes represents an unresolved puzzle in evolutionary biology. Current research suggests that eukaryotes evolved from a merger between a host of archaeal descent and an alphaproteobacterial endosymbiont. The discovery of the Asgard archaea, a proposed archaeal superphylum that includes Lokiarchaeota, Thorarchaeota, Odinarchaeota and Heimdallarchaeota suggested to comprise the closest archaeal relatives of eukaryotes, has helped to elucidate the identity of the putative archaeal host. Whereas Lokiarchaeota are assumed to employ a hydrogen-dependent metabolism, little is known about the metabolic potential of other members of the Asgard superphylum. We infer the central metabolic pathways of Asgard archaea using comparative genomics and phylogenetics to be able to refine current models for the origin of eukaryotes. Our analyses indicate that Thorarchaeota and Lokiarchaeota encode proteins necessary for carbon fixation via the Wood-Ljungdahl pathway and for obtaining reducing equivalents from organic substrates. By contrast, Heimdallarchaeum LC2 and LC3 genomes encode enzymes potentially enabling the oxidation of organic substrates using nitrate or oxygen as electron acceptors. The gene repertoire of Heimdallarchaeum AB125 and Odinarchaeum indicates that these organisms can ferment organic substrates and conserve energy by coupling ferredoxin reoxidation to respiratory proton reduction. Altogether, our genome analyses suggest that Asgard representatives are primarily organoheterotrophs with variable capacity for hydrogen consumption and production. On this basis, we propose the 'reverse flow model', an updated symbiogenetic model for the origin of eukaryotes that involves electron or hydrogen flow from an organoheterotrophic archaeal host to a bacterial symbiont.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.002 | 0.003 |
| Open science | 0.002 | 0.001 |
| Research integrity | 0.002 | 0.002 |
| Insufficient payload (model declined to judge) | 0.004 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".