The genomic basis of reproductive and migratory behaviour in a polymorphic salmonid
Bibliographic record
Abstract
Recent ecotypic differentiation provides unique opportunities to investigate the genomic basis and architecture of local adaptation, while offering insights into how species form and persist. Sockeye salmon (Oncorhynchus nerka) exhibit migratory and resident ("kokanee") ecotypes, which are further distinguished into shore-spawning and stream-spawning reproductive ecotypes. Here, we analysed 36 sockeye (stream-spawning) and kokanee (stream- and shore-spawning) genomes from a system where they co-occur and have recent common ancestry (Okanagan Lake/River in British Columbia, Canada) to investigate the genomic basis of reproductive and migratory behaviour. Examination of the genomic landscape of differentiation, differences in allele frequencies and genotype-phenotype associations revealed three main blocks of sequence differentiation on chromosomes 7, 12 and 20, associated with migratory behaviour, spawning location and spawning timing. Structural variants identified in these same areas suggest they could contribute to ecotypic differentiation directly as causal variants or via maintenance of their genomic architecture through recombination suppression mechanisms. Genes in these regions were related to spatial memory and swimming endurance (SYNGAP, TPM3), as well as eye and brain development (including SIX6), potentially associated with differences in migratory behaviour and visual habitats across spawning locations, respectively. Additional genes (GREB1L, ROCK1) identified here have been associated with timing of migration in other salmonids and could explain variation in timing of O. nerka spawning. Together, these results based on the joint analysis of sequence and structural variation represent a significant advance in our understanding of the genomic landscape of ecotypic differentiation at different stages in the speciation continuum.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".