Genomic resources for comparative analyses of avian obligate brood parasitism
Bibliographic record
Abstract
Examples of convergent evolution, wherein distantly related organisms evolve similar traits, including behaviors, underscore the adaptive power of natural selection. In birds, obligate brood parasitism, and the associated loss of parental care behaviors, has evolved independently in seven different lineages, though little is known about the genetic basis of the complex suite of traits associated with this rare life history strategy. We generated genome assemblies for ten brood parasitic species plus eight species representatives of their parental/nesting outgroups. This includes nine long-read chromosome-level assemblies, with scaffold N50 sizes ranging from 38.1 to 72.6 MB, and gene representation completeness measures >97%. Leveraging this new catalog of avian genomes, we constructed clade-level alignments that reveal variation in chromosomal synteny, provide first-time or improved annotations of protein-coding and non-coding genes, and define cross-species ortholog reference sets. We also refine estimates for the timing of the seven independent origins of brood parasitism, ranging from recent events such as 1.6-4.5 million years ago in Molothrus cowbirds to much earlier origins over 30 million years ago in two of the three cuckoo lineages. These genomic resources lay the foundation for investigating the genetic and genomic underpinnings of brood parasitism, including the loss of parental care, shifts in mating systems, perhaps resulting in heightened sperm competition, elevated annual fecundity, improved spatial cognition related to nest-finding, and the diverse adaptations shaped by intense coevolution with host species.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.006 |
| Meta-epidemiology (narrow) | 0.002 | 0.002 |
| Meta-epidemiology (broad) | 0.002 | 0.002 |
| Bibliometrics | 0.007 | 0.013 |
| Science and technology studies | 0.002 | 0.000 |
| Scholarly communication | 0.002 | 0.001 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.001 | 0.002 |
| Insufficient payload (model declined to judge) | 0.028 | 0.019 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".