Phylogeny of endocytic components yields insight into the process of nonendosymbiotic organelle evolution
Bibliographic record
Abstract
The process by which some eukaryotic organelles, for example the endomembrane system, evolved without endosymbiotic input remains poorly understood. This problem largely arises because many major cellular systems predate the last common eukaryotic ancestor (LCEA) and thus do not provide examples of organellogenesis in progress. A model is emerging whereby gene duplication and divergence of multiple "specificity-" or "identity-" encoding proteins for the various endomembranous organelles produced the diversity of nonendosymbiotically derived cellular compartments present in modern eukaryotes. To address this possibility, we analyzed three molecular components of the endocytic membrane-trafficking machinery. Phylogenetic analyses of the endocytic syntaxins, Rab 5, and the beta-adaptins each reveal a pattern of ancestral, undifferentiated endocytic homologues in the LCEA. Subsequently, these undifferentiated progenitors independently duplicated in widely divergent lineages, convergently producing components with similar endocytic roles, e.g., beta1 and beta2-adaptin. In contrast, beta3, beta4, and all other adaptin complex subunits, as well as paralogues of the syntaxins and Rabs specific for the other membrane-trafficking organelles, all evolved before the LCEA. Thus, the process giving rise to the differentiated organelles of the endocytic system appears to have been interrupted by the major speciation event that produced the extant eukaryotic lineages. These results suggest that although many endocytic components evolved before the LCEA, other major features evolved independently and convergently after diversification into the primary eukaryotic supergroups. This finding provides an example of a basic cellular system that was simpler in the LCEA than in many extant eukaryotes and yields insight into nonendosymbiotic organelle evolution.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".