A spatially explicit method for evaluating accuracy of species distribution models
Bibliographic record
Abstract
Abstract Aim Models predicting the spatial distribution of animals are increasingly used in wildlife management and conservation planning. There is growing recognition that common methods of evaluating species distribution model (SDM) accuracy, as a global overall value of predictive ability, could be enhanced by spatially evaluating the model thereby identifying local areas of relative predictive strength and weakness. Current methods of spatial SDM model assessment focus on applying local measures of spatial autocorrelation to SDM residuals, which require quantitative model outputs. However, SDM outputs are often probabilistic (relative probability of species occurrence) or categorical (species present or absent). The goal of this paper was to develop a new method, using a conditional randomization technique, which can be applied to directly spatially evaluate probabilistic and categorical SDMs. Location Eastern slopes, Rocky Mountains, Alberta, Canada. Methods We used predictions from seasonal grizzly bear ( Ursus arctos ) resource selection functions (RSF) models to demonstrate our spatial evaluation technique. Local test statistics computed from bear telemetry locations were used to identify areas where bears were located more frequently than predicted. We evaluated the spatial pattern of model inaccuracies using a measure of spatial autocorrelation, local Moran’s I . Results We found the model to have non‐stationary patterns in accuracy, with clusters of inaccuracies located in central habitat areas. Model inaccuracies varied seasonally, with the summer model performing the best and the least error in areas with high RSF values. The landscape characteristics associated with model inaccuracies were examined, and possible factors contributing to RSF error were identified. Main conclusions The presented method complements existing spatial approaches to model error assessment as it can be used with probabilistic and categorical model output, which is typical for SDMs. We recommend that SDM accuracy assessments be done spatially and resulting accuracy maps included in model metadata.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".