A Canadian oat genomic selection study incorporating genetic and environmental information
Notice bibliographique
Résumé
Oat (Avena sativa L.) is an important crop in Canada that has been seeded on an average of 3.3 million acres over the past five years. It is considered a healthy cereal due to the presence of beta-glucan in the grain, which been shown to reduce the risk of heart disease, as well as being a good source of protein that is rich in globulins. Identifying new breeding strategies that can improve breeding efficiency in oat is important for future progress in this crop. To this end, genomic and environmental factors, along with their interactions, were examined to determine what contributed to variation in important oat traits. This information was then used to develop genomic selection (GS) models that can be used in oat breeding programs.\nIn the first study, 305 elite oat breeding lines grown in the Western Cooperative Oat Registration Trial (WCORT) from 2002 to 2014 were used to investigate important factors for genomic selection model building. The influence of phenotypic data, genotyping platforms, statistical model, marker density, population structure, training population size and trait heritability were assessed. It was determined that the machine learning model Support Vector Machine and the additive linear model rr-BLUP offered the best overall prediction accuracies. Prediction accuracy increased when using the iSelect Oat 6K SNP chip, as the marker number increased, with larger training population size and with traits that were more heritable.\nIn the second study, environmental and correlated agronomic variables, along with their inter-relationships, that contributed to variation in yield and grain β-glucan content in oat lines was investigated. A hypothesized structural equation model (SEM) that included variables related to environmental and phenotypic traits was created and tested against observed yield data. Significant paths were identified to explain yield variation (59%-76%) among the three oat varieties. A similar approach was taken for β-glucan in which significant paths were found which explained 16%-41% of the variation in β-glucan. Results from this study suggest that a longer period to heading and maturity, and a taller stature were the three phenotypic traits that most positively influence yield. Limited precipitation before maturity, high temperatures during heading and grain filling were the three environmental variables that contributed to decreased yield. Precipitation and July temperature were the two most important environmental variables that influenced β-glucan, while maturity was the most important trait affecting β-glucan, although the direction of effect for maturity varied by oat variety.\nIn the third study, additional information was added into the previous GS models to determine if prediction could be improved. Genotype, environment and their interaction were used to conduct genomic selection for yield. Four mega-environments were identified from Ward’s hierarchical clustering using the significant environmental variables identified in the second study. It was found that using individual locations to represent environment provided more accuracy compared to using mega-environments. The reaction norm model was also tested which allowed significant environmental variables to be incorporated as a covariance matrix in the model. Including an environmental covariance matrix and interaction terms increased prediction accuracy compared to models with only genotype main effects. Multiple trait GS did not provide better prediction accuracy for most the traits.\n In the final study, GS was used to predict the GEBVs of two populations, a biparental derived population and a population consisting of elite breeding lines from several different breeding programs. Higher predication accuracy was found in the elite breeding line population which was likely due to the closer genetic relationship between it and the training population. Finally, random selection and genomic selection were compared in the two populations. Genomic selection out-performed random selection in the elite breeding population, but not in the bi-parental population. Again, the poor performance of GS in the bi-parental population was best explained by the unrelatedness between it and the training population.\nTaken together, these studies provided deeper insight into how GS could be applied in oat breeding programs.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,001 | 0,001 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,001 |
| Bibliométrie | 0,001 | 0,002 |
| Études des sciences et des technologies | 0,002 | 0,001 |
| Communication savante | 0,001 | 0,000 |
| Science ouverte | 0,001 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,001 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».