Historical Selection, Adaptation Signatures, and Ambiguity of Introgressions in Wheat
Bibliographic record
Abstract
Wheat was one of the crops domesticated in the Fertile Crescent region approximately 10,000 years ago. Despite undergoing recent polyploidization, hull-to-free-thresh transition events, and domestication bottlenecks, wheat is now grown in over 130 countries and accounts for a quarter of the world's cereal production. The main reason for its widespread success is its broad genetic diversity that allows it to thrive in different environments. To trace historical selection and hybridization signatures, genome scans were performed on two datasets: approximately 113K SNPs from 921 predominantly bread wheat accessions and approximately 110K SNPs from about 400 wheat accessions representing all ploidy levels. To identify environmental factors associated with the loci, a genome-environment association (GEA) was also performed. The genome scans on both datasets identified a highly differentiated region on chromosome 4A where accessions in the first dataset were dichotomized into a group (n = 691), comprising nearly all cultivars, wild emmer, and most landraces, and a second group (n = 230), dominated by landraces and spelt accessions. The grouping of cultivars is likely linked to their potential ancestor, bread wheat cv. Norin-10. The 4A region harbored important genes involved in adaptations to environmental conditions. The GEA detected loci associated with latitude and temperature. The genetic signatures detected in this study provide insight into the historical selection and hybridization events in the wheat genome that shaped its current genetic structure and facilitated its success in a wide spectrum of environmental conditions. The genome scans and GEA approaches applied in this study can help in screening the germplasm housed in gene banks for breeding, and for conservation purposes.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.002 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".