Use of deep convolutional neural networks to predict Parkinson’s disease progression from DaTscan SPECT images
Notice bibliographique
Résumé
29 Objectives: Brain SPECT imaging with Ioflupane I123 (DaTscan) is clinically used to measure the striatal dopamine transporter levels and aid the diagnosis of Parkinson’s disease (PD). It has been shown that the mean tracer binding ratios (BRs) in manually-defined regions of interest (ROIs) may also aid in predicting the rate of disease progression, e.g. cognitive decline (Caspell-Garcia, 2017) [1]. However, the mean BR may neglect other disease-relevant image features, such as shapes and texture of the tracer distribution. Recently, deep learning has gained popularity as a promising approach that can lead to fundamental advances in imaging-based biomarkers (Vieira, 2017) [2]. Our objective was to test whether deep convolutional neural networks (CNN) can be trained to extract salient information from DaTscan images - features that are invariant to translation, rotation, and blurring. We train the CNNs that operate on voxels within simple bounding box (BB) ROIs to predict cognitive decline in de-novo PD patients. We compare the prediction accuracy between the CNN and conventional mean BRs within manually-placed ROIs (via a logistic classifier). Methods: Baseline DaTscan images (91x109x91 voxels with size 2.0 mm) for 370 de-novo PD subjects as well as mean BR values in the putamen and caudate were obtained from the PPMI (www.ppmi-info.org). For CNN training, the images were cropped to a BB (32x48x16 voxels) that encompassed the striatum. The predicted metric was the clinical diagnosis of cognitive impairment at 2-year follow-up: 305 subjects were not cognitively impaired (NCI), and 65 were impaired (CI). The subjects were split into stratified, equally-sized train and test sets with approximately equal NCI to CI ratio (Fig. 1). The train BRs were used to fit a logistic classifier. The train BB images were augmented for CNN training, by adding random rotation, translation and blurring, yielding 50,000 images. The test set was used to measure the receiver operating characteristic (ROC) area-under the curve (AUC) of the logistic and CNN classifiers. A CNN with 3 convolution layers (Fig. 2) was implemented using the TensorFlow library (https://www.tensorflow.org). Each convolution layer was followed by a Rectifying Linear Unit (ReLU) activation and a max-pooling layer. The cross-entropy was the optimized loss function. Each training iteration included 50 augmented BB images. Results: The train and test loss functions plotted against the iteration number (Fig. 3A) demonstrate the CNN learning process: the test loss was reduced in the process of train loss optimization. The AUC of the CNN increased with iteration number (Fig. 3B), and stabilized after approximately 400 iterations (~20,000 train BB images). Multiple training repetitions produced similar results. The mean CNN AUC from iterations 400-600 was 0.69±0.02, outperforming the AUC of 0.63 for conventional mean BR values (Fig. 3C). When age was included as a covariate, the AUC of both classifiers was 0.74, similar to previous findings with mean BR (Schrag, 2017) [3]. Conclusions: Deep CNNs can be trained to use DaTscan SPECT images for prediction of cognitive impairment in PD subjects. The network learned disease-related image features that are invariant to rotation, translation, and blurring. The AUC comparison suggests that CNNs may predict disease progression better than mean BRs. Thus, in conjunction with the simplified BB ROIs, deep CNN offer a fully-automated approach to DaTscan image analysis. Since the image features learned by the CNN are rotation and translation-invariant, the DaTscan images do not need to be realigned, and labor-intensive manual ROI placement is not required.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction distillée sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.
Scores Codex et Gemma par catégorie
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,000 | 0,000 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,000 | 0,000 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,000 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,001 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».