The phylogeny and global biogeography of Primulaceae based on high-throughput DNA sequence data
Bibliographic record
Abstract
The angiosperm family Primulaceae is morphologically diverse and distributed nearly worldwide. However, phylogenetic uncertainty has obstructed the identification of major morphological and biogeographic transitions within the clade. We used target capture sequencing with the Angiosperms353 probes, taxon-sampling encompassing nearly all genera of the family, tree-based sequence curation, and multiple phylogenetic approaches to investigate the major clades of Primulaceae and their relationship to other Ericales. We generated dated phylogenetic trees and conducted broad-scale biogeographic analyses as well as stochastic character mapping of growth habit. We show that Ardisia, a pantropical genus and the largest in the family, is not monophyletic, with at least 19 smaller genera nested within it. Neotropical members of Ardisia and several smaller genera form a clade, an ancestor of which arrived in the Neotropics and began diversifying about 20 Ma. This Neotropical clade is most closely related to Elingamita and Tapeinosperma, which are most diverse on islands of the Pacific. Both Androsace and Primula are non-monophyletic by the inclusion of smaller genera. Ancestral state reconstructions revealed that there have either been parallel transitions to an herbaceous habit in Primuloideae, Samolus, and at least three lineages of Myrsinoideae, or a common ancestor of nearly all Primulaceae was herbaceous. Our results provide a robust estimate of phylogenetic relationships across Primulaceae and show that a revised classification of Myrsinoideae and several other clades within the family is necessary to render all genera monophyletic.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".