A new high-quality genome assembly and annotation for the threatened Florida Scrub-Jay ( <i>Aphelocoma coerulescens</i> )
Bibliographic record
Abstract
The Florida Scrub-Jay (Aphelocoma coerulescens), a Federally Threatened, cooperatively-breeding bird, is an emerging model system in evolutionary biology and ecology. Extensive individual-based monitoring and genetic sampling for decades has yielded a wealth of data, allowing for the detailed study of social behavior, demography, and population genetics of this natural population. Here, we report a linkage map and a chromosome-level genome assembly and annotation for a female Florida Scrub-Jay made with long-read sequencing technology, chromatin conformation data, and the linkage map. We constructed a linkage map comprising 4,468 SNPs that had 34 linkage groups and a total sex-averaged autosomal genetic map length of 2446.78 cM. The new genome assembly is 1.33 Gb in length, consisting of 33 complete or near-complete autosomes and the sex chromosomes (ZW). This highly contiguous assembly has an NG50 of 68 Mb and a Benchmarking Universal Single-Copy Orthologs (BUSCO) completeness score of 97.1% with respect to the Aves database. The annotated gene set has a BUSCO transcriptome completeness score of 95.5% and 17,964 identified protein-coding genes, 92.5% of which have associated functional annotations. This new, high-quality genome assembly and linkage map of the Florida Scrub-Jay provides valuable tools for future research into the evolutionary dynamics of small, natural populations of conservation concern.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.004 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".