A Comparison of SimCLR and SwAV Contrastive Self-Supervised Learning Models For Landslide Detection
Notice bibliographique
Résumé
Deep Learning (DL) algorithms have demonstrated superior efficacy compared to traditional Machine Learning (ML) methods in the realm of landslide detection through the analysis of Remote Sensing (RS) imagery. However, their performance is notably contingent upon the quantity of manual annotations utilized during the training process. This investigation delves into the utilization of two distinct Self-Supervised Learning (SSL) models, specifically the Simple Framework for Contrastive Learning of Visual Representations (SimCLR) and Swapping Assignments between multiple Views (SwAV). These models were adapted and enhanced for downstream tasks, particularly in the domain of landslide detection. To train the SSL models, the Landslide4Sense competition dataset was employed, consisting of 3799 training patches, 245 validation patches, and 800 testing patches generated from Sentinel-2 images acquired from diverse regions worldwide. During the training of SimCLR and SwAV models, only the training patches were utilized, with a series of data augmentations applied to the input dataset based on each model's architecture. Both models employed ResNet-50 as the encoder.For the downstream task of landslide detection, a custom U-Net model was developed. The trained ResNet-50 served as the encoder, and during fine-tuning, only the decoder part was permitted to be trained while the encoder remained frozen. During the fine-tuning process, subsets comprising 1% and 10% of labeled data from the training dataset were randomly selected to train the model, and predictions were exclusively conducted on the testing data. While a conventional supervised ResU-Net model, which was trained on all labeled training datasets, attained an F1 score of 72%, the SSL models achieved F1 scores of 64% and 71% with 1% labeled data, and 68% and 76% with 10% labeled data for SimCLR and SwAV, respectively. In addition, comparisons were conducted with all supervised reference models in the Landslide4Sense competition, revealing that SwAV, with 10% labeled data, outperformed all models, surpassing their top model by 4%. This study underscores the potential of SSL techniques in the segmentation and classification of RS images for natural hazard mapping, particularly in scenarios where labeled data is not available or is limited.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,004 | 0,004 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,001 | 0,001 |
| Études des sciences et des technologies | 0,000 | 0,001 |
| Communication savante | 0,001 | 0,002 |
| Science ouverte | 0,003 | 0,001 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,001 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».