Sampling strategies for assessing lameness, injuries, and body condition score on dairy farms
Notice bibliographique
Résumé
Our objective was to evaluate how sampling strategies (i.e., how many cows to sample and which animals to include) used in 4 dairy cattle welfare assessment programs affect the classification of dairy farms relative to thresholds of acceptability for animal-based measures. We predicted that classification performance would improve when more cows were sampled and when selecting from all lactating cows versus when some pens were excluded. On 38 freestall farms, we assessed all 12,375 cows for lameness, injuries on the tarsal (hock) and carpal joints, and body condition score and calculated the farm-level prevalence for each measure. Based on approaches used in the industry, we evaluated 6 sampling strategies generated using formulas with precision (d) of 15, 10, or 5% applied to either a single high-producing pen or all lactating cows; an additional sample was included with d = 10% applied to the entire herd, selecting lactating cows in proportion to their representation in the herd. For each sampling strategy, cow records were selected randomly (in 10,000 replicates) to calculate prevalence. The strategy of assessing all cows in the high-producing pen was also compared. Farms were classified as meeting (below) or failing to meet (above) thresholds of ≤15% moderate lameness; ≤20% moderate carpal or hock injuries; <10, <5, and ≤1% severe lameness; or injuries on the carpus or hock; and <5, <3, <1, or 0% thin cows. For each measure and threshold, we calculated pooled percent agreement, kappa, sensitivity, specificity, and positive and negative predictive value for each sampling strategy using true prevalence as the gold standard for herd classification. Across measures and thresholds, classification performance increased with the number of cows sampled [i.e., when narrower precision values (d = 5 vs. 10 vs. 15%) were used in the sample size calculation]. Because narrower precision values can dramatically increase sample size, assessment programs may need to consider both feasibility and the degree of misclassification they will accept. Applying the formula directly to lactating cows performed better than applying it to the entire herd and then selecting lactating cows in proportion to their representation in the herd. Farm classifications were similar whether cows in the hospital pen were included or excluded from the sample. Selecting all cows from the high-producing pen resulted in classifications similar to when including all lactating cows, suggesting that assessing cows from the high-producing pen may serve as an acceptable proxy for all lactating cows on the farm.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,019 | 0,030 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,000 |
| Bibliométrie | 0,001 | 0,001 |
| Études des sciences et des technologies | 0,001 | 0,000 |
| Communication savante | 0,001 | 0,001 |
| Science ouverte | 0,001 | 0,001 |
| Intégrité de la recherche | 0,001 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,000 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».