Welfare-improving enrichments greatly reduce hens’ startle responses, despite little change in judgment bias
Bibliographic record
Abstract
Responses to ambiguous and aversive stimuli (e.g. via tests of judgment bias and measures of startle amplitude) can indicate mammals' affective states. We hypothesised that such findings generalize to birds, and that these two responses co-vary (since both involve stimulus evaluation). To validate startle reflexes (involuntary responses to sudden aversive stimuli) and responses in a judgment bias task as indicators of avian affective state, we differentially housed hens with or without preferred enrichments assumed to improve mood (in a crossover design). To control for personality, we first measured hens' baseline exploration levels. To infer judgment bias, control and enriched hens were trained to discriminate between white and dark grey cues (associated with reward and punishment, respectively), and then probed with intermediate shades of grey. For startle reflexes, forceplates assessed responses to a light flash. Judgment bias was only partially validated: Exploratory hens showed more 'optimism' when enriched, but Non-exploratory hens did not. Across all birds, however, startle amplitudes were dramatically reduced by enrichment (albeit more strongly in Exploratory subjects): the first evidence that avian startle is affectively modulated. Startle and judgment biases did not co-vary, suggesting different underlying mechanisms. Of the two measures, startle reflexes thus seem most sensitive to avian affective state.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".