How smart was <scp> <i>T. rex</i> </scp> ? Testing claims of exceptional cognition in dinosaurs and the application of neuron count estimates in palaeontological research
Bibliographic record
Abstract
Recent years have seen increasing scientific interest in whether neuron counts can act as correlates of diverse biological phenomena. Lately, Herculano-Houzel (2023) argued that fossil endocasts and comparative neurological data from extant sauropsids allow to reconstruct telencephalic neuron counts in Mesozoic dinosaurs and pterosaurs, which might act as proxies for behaviors and life history traits in these animals. According to this analysis, large theropods such as Tyrannosaurus rex were long-lived, exceptionally intelligent animals equipped with "macaque- or baboon-like cognition", whereas sauropods and most ornithischian dinosaurs would have displayed significantly smaller brains and an ectothermic physiology. Besides challenging established views on Mesozoic dinosaur biology, these claims raise questions on whether neuron count estimates could benefit research on fossil animals in general. Here, we address these findings by revisiting Herculano-Houzel's (2023) work, identifying several crucial shortcomings regarding analysis and interpretation. We present revised estimates of encephalization and telencephalic neuron counts in dinosaurs, which we derive from phylogenetically informed modeling and an amended dataset of endocranial measurements. For large-bodied theropods in particular, we recover significantly lower neuron counts than previously proposed. Furthermore, we review the suitability of neurological variables such as neuron numbers and relative brain size to predict cognitive complexity, metabolic rate and life history traits in dinosaurs, coming to the conclusion that they are flawed proxies for these biological phenomena. Instead of relying on such neurological estimates when reconstructing Mesozoic dinosaur biology, we argue that integrative studies are needed to approach this complex subject.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".