Standardization of milk infrared spectra for the retroactive application of calibration models
Notice bibliographique
Résumé
The objective of this study was to standardize the infrared spectra obtained over time and across 2 milk laboratories of Canada to create a uniform historical database and allow (1) the retroactive application of calibration models for prediction of fine milk composition; and (2) the direct use of spectral information for the development of indicators of animal health and efficiency. Spectral variation across laboratories and over time was inspected by principal components analysis (PCA). Shifts in the PCA scores were detected over time, leading to the definition of different subsets of spectra having homogeneous infrared signal. To evaluate the possibility of using common equations on spectra collected by the 2 instruments and over time, we developed a standardization (STD) method. For each subset of data having homogeneous infrared signal, a total of 99 spectra corresponding to the percentiles of the distribution of the absorbance at each wavenumber were created and used to build the STD matrices. Equations predicting contents of saturated fatty acids, short-chain fatty acids, and C18:0 were created and applied on different subsets of spectra, before and after STD. After STD, bias and root mean squared error of prediction decreased by 66% and 32%, respectively. When calibration equations were applied to the historical nonstandardized database of spectra, shifts in the predictions could be observed over time for all investigated traits. Shifts in the distribution of the predictions over time corresponded to the shifts identified by the inspection of the PCA scores. After STD, shifts in the predicted fatty acid contents were greatly reduced. Standardization reduced spectral variability between instruments and over time, allowing the merging of milk spectra data from different instruments into a common database, the retroactive use of calibrations equations, or the direct use of the spectral data without restrictions.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,016 | 0,029 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,001 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,002 | 0,003 |
| Études des sciences et des technologies | 0,001 | 0,001 |
| Communication savante | 0,002 | 0,001 |
| Science ouverte | 0,001 | 0,001 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,001 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».