Comprehensive phylogenomic time tree of bryophytes reveals deep relationships and uncovers gene incongruences in the last 500 million years of diversification
Bibliographic record
Abstract
PREMISE: Bryophytes form a major component of terrestrial plant biomass, structuring ecological communities in all biomes. Our understanding of the evolutionary history of hornworts, liverworts, and mosses has been significantly reshaped by inferences from molecular data, which have highlighted extensive homoplasy in various traits and repeated bursts of diversification. However, the timing of key events in the phylogeny, patterns, and processes of diversification across bryophytes remain unclear. METHODS: Using the GoFlag probe set, we sequenced 405 exons representing 228 nuclear genes for 531 species from 52 of the 54 orders of bryophytes. We inferred the species phylogeny from gene tree analyses using concatenated and coalescence approaches, assessed gene conflict, and estimated the timing of divergences based on 29 fossil calibrations. RESULTS: The phylogeny resolves many relationships across the bryophytes, enabling us to resurrect five liverwort orders and recognize three more and propose 10 new orders of mosses. Most orders originated in the Jurassic and diversified in the Cretaceous or later. The phylogenomic data also highlight topological conflict in parts of the tree, suggesting complex processes of diversification that cannot be adequately captured in a single gene-tree topology. CONCLUSIONS: We sampled hundreds of loci across a broad phylogenetic spectrum spanning at least 450 Ma of evolution; these data resolved many of the critical nodes of the diversification of bryophytes. The data also highlight the need to explore the mechanisms underlying the phylogenetic ambiguity at specific nodes. The phylogenomic data provide an expandable framework toward reconstructing a comprehensive phylogeny of this important group of plants.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".