Spatial autocorrelation in fitness affects the estimation of natural selection in the wild
Bibliographic record
Abstract
Summary Natural selection is typically estimated in the wild using Lande and Arnold's multiple regression approach. Despite its utility for evolutionary ecologists, this method is subject to the classical assumptions of multiple regressions, which could result in potential analytical problems. In particular, spatial autocorrelation in fitness violates the assumption of residuals independence. Although widespread in the wild, the consequences of this effect have yet to be investigated in the context of Lande and Arnold's regression and resulting selection estimation. Here we first described four spatially explicit models that allow to control for spatial autocorrelation in residuals of the Lande and Arnold's regression: a generalized least square ( GLS ) model with a distance‐based exponential covariance function, two simultaneous autoregressive models ( SAR , the lagged‐response model ( SAR ‐lag) and the spatial error model ( SAR ‐err)) and a 5‐step procedure using the principal coordinates of neighbour matrices ( PCNM ) method based on the extraction of spatial descriptors. We then compared the four spatially explicit models of selection to non‐spatial models for three life‐history traits recorded over 6 years in a wild blue tit ( Cyanistes caeruleus ) population. We also compared the performance of the four spatially explicit models of selection using a simulation approach. Our analyses revealed strong spatial autocorrelation in residuals of selection models, which was completely described by the two SAR and the PCNM models, while only partially described by the GLS model. The magnitude of selection gradients and differentials decreased systematically in the 4 spatially explicit models while the degree of fit of these models increased (except for the GLS model). Moreover, we showed using simulations that the selection coefficients extracted from the SAR ‐lag model were systematically biased compared to those extracted from the GLS , SAR ‐err and PCNM models. We hereby showed that spatial autocorrelation in fitness can severely affect selection differentials and gradients, even at a relatively small spatial scale. By using geostatistical models such as PCNM or SAR ‐err models, it is possible to control for this spatial autocorrelation. Finally, since spatial autocorrelation is closely linked to spatial environmental variation, this approach can also be used to explore environmental components of covariance between fitness and traits.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".