Phylogeny, biogeography and diversification of the mining bee family Andrenidae
Bibliographic record
Abstract
Abstract The mining bees (Andrenidae) are a major bee family of over 3000 described species with a nearly global distribution. They are a particularly significant component of northern temperate ecosystems and are critical pollinators in natural and agricultural settings. Despite their ecological and evolutionary significance, our knowledge of the evolutionary history of Andrenidae is sparse and insufficient to characterize their spatiotemporal origin and phylogenetic relationships. This limits our ability to understand the diversification dynamics that led to the second most species‐rich genus of all bees, Andrena Fabricius, and the most species‐rich North American genus, Perdita Smith. Here, we develop a comprehensive genomic dataset of 195 species of Andrenidae, including all major lineages, to illuminate the evolutionary history of the family. Using fossil‐informed divergence time estimates, we characterize macroevolutionary dynamics, incorporate paleoclimatic information, and present our findings in the context of diversification rate estimates for all other bee tribes. We found that diversification rates of Andrenidae steeply increased over the past 15 million years, particularly in the genera Andrena and Perdita . This suggests that these two groups and the brood parasites of the genus Nomada Scopoli (Apidae), which are the primary cleptoparasitic counterparts of Andrena , are similar in age and represent the fastest diversifying lineages of all bees. Using our newly developed time frame of andrenid evolution, we estimate a late Cretaceous origin in South America for the family and reconstruct the past dispersal events that led to its present‐day distribution.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".