Bibliographic record
Abstract
Background: Deep network models are relatively successful at predicting both performance and neural activations on object recognition tasks. However, recent work suggests that these models rely largely on local features rather than global shape, whereas humans, while sensitive to local 2D shape (curvature), are easily able to discriminate natural shapes from synthetic shapes with matched curvature statistics. Here we assess two alternative shape models that could account for this human sensitivity to 2D shape beyond local curvature: 1) Pooling – shape information is pooled over a collection of independently-coded fragments or parts; 2) Configural – the representation depends on the arrangement of these parts over the entire shape. Method: We employed a dataset of 2D animal shapes approximated as 120-segment outline polygons and local ‘metamers’ – closed contours that match the local curvature statistics of the animal shapes. In a two-interval task, five observers discriminated between a stimulus containing only animal contour fragments and a second stimulus containing only metamer fragments, while the length of the fragments was varied from 2 segments (local) to 120 segments (global). There were two conditions: 1. A single fragment displayed centrally. 2. Multiple fragments displayed within a 7.5 deg circular window. The number of fragments was selected to yield a total of 120 turning angles, matching the full-shape condition. Results: For both single- and multi-fragment conditions, performance rises from chance to near 100% as fragment length increases from 2 to 120, reflecting human sensitivity to 2D shape beyond local curvature. Interestingly, there is little difference in the psychometric functions for the single- and multi-fragment conditions (75%-correct thresholds of 24 +/- 7 vs 18 +/- 6 segments), indicating very little pooling across fragments. This suggests that human shape perception is highly configural, posing a challenge to recent deep learning accounts of object coding.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".