Leveraging Open‐Source Geographic Databases to Enhance the Representation of Landscape Heterogeneity in Ecological Models
Notice bibliographique
Résumé
Wildlife abundance and movement are strongly impacted by landscape heterogeneity, especially in cities which are among the world's most heterogeneous landscapes. Nonetheless, current global land cover maps, which are used as a basis for large-scale spatial ecological modeling, represent urban areas as a single, homogeneous, class. This often requires urban ecologists to rely on geographic resources from local governments, which are not comparable between cities and are not available in underserved countries, limiting the spatial scale at which urban conservation issues can be tackled. The recent expansion of community-based geographic databases, for example, OpenStreetMap (OSM), represents an opportunity for ecologists to generate large-scale maps geared toward their specific research needs. However, computational differences in language and format, and the high diversity of information within, limit the access to these data. We provide a framework, using R, to extract geographic features from the OSM database, classify, and integrate them into global land cover maps. The framework includes an exhaustive list of OSM features describing urban and peri-urban landscapes and is validated by quantifying the completeness of the OSM features characterized, and the accuracy of its final output in 34 cities in North America. We portray its application as the basis for generating landscape variables for ecological analysis by using the OSM-enhanced map to generate an urbanization index, and subsequently analyze the spatial occupancy of six mammals throughout Chicago, Illinois, USA. The OSM features characterized had high completeness values for impervious land cover classes (50%-100%). The final output, the OSM-enhance map, provided an 89% accurate representation of the landscape at 30m resolution. The OSM-derived urbanization index outperformed other global spatial data layers in the spatial occupancy analysis and concurred with previously seen local response trends, whereby lagomorphs and squirrels responded positively to urbanization, while skunks, raccoons, opossums, and deer responded negatively. This study provides a roadmap for ecologists to leverage the fine resolution of open-source geographic databases and apply it to spatial modeling by generating research-specific landscape variables. As our occupancy results show, using context-specific maps can improve modeling outputs and reduce uncertainty, especially when trying to understand anthropogenic impacts on wildlife populations.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,006 | 0,035 |
| Méta-épidémiologie (sens strict) | 0,002 | 0,001 |
| Méta-épidémiologie (sens large) | 0,001 | 0,003 |
| Bibliométrie | 0,006 | 0,008 |
| Études des sciences et des technologies | 0,001 | 0,001 |
| Communication savante | 0,005 | 0,005 |
| Science ouverte | 0,004 | 0,006 |
| Intégrité de la recherche | 0,001 | 0,002 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,003 | 0,002 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».