29Prognostic safety of automatic cancellation of rest myocardial perfusion scan by machine learning: a report from multicenter REFINE SPECT registry of new generation SPECT
Notice bibliographique
Résumé
Abstract Background We aimed to develop a machine learning (ML) computer score derived from stress imaging and clinical data, which indicates if the rest scan could be automatically and safely canceled in the routine stress/rest myocardial perfusion SPECT (MPS). Methods A total of 20414 stress/rest cases from the REFINE SPECT registry collected from 5 sites in 3 countries with Tc-99m-based MPS images, clinical data, and clinical follow-up were included in the study. All images were automatically processed at our Medical Center. The automatically generated myocardial contours were checked by experienced technologists. In total, 93 variables (26 clinical, 17 stress-test, and 50 stress-imaging variables) were used to build a LogitBoost model for prediction of adverse events (AE), including coronary revascularization, death, myocardial infarction, and unstable angina. 10-fold cross-validation was performed to separate test from validation data for the assessment of ML. The overall ML predictive performance was compared to quantitative (stress total perfusion deficit [TPD]) by the area under the receiver operating characteristic curves (AUC). ML cut-off (ML1) to simulate the decision of cancellation of the rest scan was set to result in the same % of normal scans as these determined by the normal clinical reader diagnosis on a 4-point scale in the whole population, or the same % of scans with visual summed stress scores (SSS) = 0 in the subpopulation with available SSS. A second ML cutoff (ML2) was established to achieve a 1% annual risk of AE. The annual risk of AE of the normal ML score was compared with normal clinical diagnosis and with the finding of SSS = 0. Results The mean follow-up interval was 4.7±1.5 years. Overall, 3542 AE were observed (3.7% annual risk). The AUC for AE was higher for ML (0.780±0.005) than for stress TPD (0.698±0.006) (p<0.001). Normal clinical diagnosis was reported in 60% cases. In 70% (14242 scans) with available segmental scores, 53% had SSS=0. ML1 and ML2 thresholds were compared with normal visual diagnosis and with SSS = 0 for AE (Figure). ML1 achieved a lower annual risk (1.5%) than normal clinical diagnosis (2.1%) or SSS = 0 (1.6% versus 2.3%) (p<0.001). The more conservative ML2 threshold with a 1% annual risk of AE resulted in a 40% canceling rate. Figure 1 Conclusion ML could be used to automatically cancel the rest MPS scan with the same proportion as using normal visual MPS reading, but with significantly lower AE rate in stress-only scans. Acknowledgement/Funding R01HL089765 from the National Heart, Lung, and Blood Institute/National Institutes of Health (NHLBI/NIH)
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,008 | 0,018 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,000 |
| Bibliométrie | 0,001 | 0,001 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,001 | 0,000 |
| Science ouverte | 0,001 | 0,001 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,001 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».