A Combined Deep Learning and Prior Knowledge Constraint Approach for Large-Scale Forest Disturbance Detection Using Time Series Remote Sensing Data
Notice bibliographique
Résumé
The scale and severity of forest disturbances across the globe are increasing due to climate change and human activities. Remote sensing analysis using time series data is a powerful approach for detecting large-scale forest disturbances and describing detailed forest dynamics. Various large-scale forest disturbance detection algorithms have been proposed, but most of them are only suitable for detecting high-magnitude forest disturbances (e.g., fire, harvest). Conversely, more continuous, subtle, and gradual lower-magnitude forest disturbances (e.g., thinning, pests, and diseases) have been subject to less focus. Deep learning (DL) can distinguish subtle differences in information within time series data, offering new opportunities to capture forest disturbances in a complete and detailed way. This study proposes an approach for analyzing forest dynamics across large areas and long time periods by combining DL time series classification and prior knowledge constraint. The approach consists of two stages: (1) an improved self-attention model used for time series classification to identify sequences with forest disturbance characteristics; (2) developed skip-disturbance recovery index (S-DRI) characterizing the temporal context, using prior knowledge constraint to identify forest disturbance years in time series with disturbance characteristics. In this study, the year of forest disturbances in five study areas located in the United States, Canada, and Poland from 2001 to 2020 was mapped. A total of 3082 manually interpreted test data with different disturbance causal agents (such as fire, harvest, conversion, hurricane, and pests) were sampled from five research areas for validation. Our approach was also evaluated against two forest disturbance benchmark datasets derived from LandTrendr and the Global Forest Change (GFC) dataset. The results demonstrate that our approach achieved an overall accuracy of 87.8%, surpassing the accuracy of LandTrendr (84.6%) and the Global Forest Change dataset (81.4%). Furthermore, our approach demonstrated lower omission rates (ranging from 10.0% to 67.4%) in detecting subtle to severe causal agents of forest disturbance, in comparison to LandTrendr (with a range of 18.0% to 81.6%) and GFC (with a range of 15.0% to 88.8%). This study, which involved mapping large-scale and long-term forest disturbance in multiple regions, revealed that our approach can be applied to new areas without a requirement for complex parameter adjustments. These results demonstrate the potential of our approach in generating comprehensive and detailed forest disturbance data, thus providing a new and effective method in this domain.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,001 | 0,002 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,001 | 0,002 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,001 | 0,001 |
| Science ouverte | 0,001 | 0,001 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,001 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».