Validation of an artificial intelligence-based method to automate Cobb angle measurement on spinal radiographs of children with adolescent idiopathic scoliosis
Notice bibliographique
Résumé
BACKGROUND: Accurately measuring the Cobb angle on radiographs is crucial for diagnosis and treatment decisions for adolescent idiopathic scoliosis (AIS). However, manual Cobb angle measurement is time-consuming and subject to measurement variation, especially for inexperienced clinicians. AIM: This study aimed to validate a novel artificial-intelligence-based (AI) algorithm that automatically measures the Cobb angle on radiographs. DESIGN: This is a retrospective cross-sectional study. SETTING: The population of patients attended the Stollery Children's Hospital in Alberta, Canada. POPULATION: Children who: 1) were diagnosed with AIS, 2) were aged between 10 and 18 years old, 3) had no prior surgery, and 4) had a radiograph out of brace, were enrolled. METHODS: A total of 330 spinal radiographs were used. Among those, 130 were used for AI model development and 200 were used for measurement validation. Automatic Cobb angle measurements were validated by comparing them with manual ones measured by a rater with 20+ years of experience. Analysis was performed using the standard error of measurement (SEM), inter-method intraclass correlation coefficient (ICC<inf>2,1</inf>), and percentage of measurements within clinical acceptance (≤5°). Subgroup analysis was conducted by severity, region, and X-ray system to identify any systematic biases. RESULTS: The AI method detected 346 of 352 manually measured curves (mean±standard deviation: 24.7±9.5°), achieving 91% (316/346) of measurements within clinical acceptance. Excellent reliability was obtained with 0.92 ICC and 0.79° SEM. Comparable performance was found throughout all subgroups, and no systematic biases in performance affecting any subgroup were discovered. The algorithm measured each radiograph approximately 18s on average which is slightly faster than the estimated measurement time of an experienced rater. Radiographs taken by the EOS X-ray system were measured more quickly on average than those taken by a conventional digital X-ray system (10s vs. 26s). CONCLUSIONS: An AI-based algorithm was developed to measure the Cobb angle automatically on radiographs and yielded reliable measurements quickly. The algorithm provides detailed images on how the angles were measured, providing interpretability that can give clinicians confidence in the measurements. CLINICAL REHABILITATION IMPACT: Employing the algorithm in practice could streamline clinical workflow and optimize measurement accuracy and speed in order to inform AIS treatment decisions.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction distillée sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.
Scores Codex et Gemma par catégorie
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,002 | 0,000 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,000 |
| Bibliométrie | 0,001 | 0,001 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,000 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,000 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».