MétaCan
Menu
Retour à la cohorte
Enregistrement W4412354059 · doi:10.1101/2025.07.10.25331257

Machine learning algorithm to predict fragility fractures and identification of important features – an explainable approach

2025· preprint· en· W4412354059 sur OpenAlexafffundabout
Sayem Borhan, Αλεξάνδρα Παπαϊωάννου, Jonathan D. Adachi, Suzanne N. Morin, David Goltzman, David A. Hanley, Claudie Berger, Lehana Thabane, Parminder Raina

Notice bibliographique

RevuemedRxiv · 2025
Typepreprint
Langueen
DomaineEngineering
ThématiqueMedical Imaging and Analysis
Établissements canadiensHamilton Health SciencesMcMaster UniversityMcGill University Health CentreSt. Joseph’s Healthcare HamiltonUniversity of CalgaryMcGill UniversityImpact
Organismes subventionnairesCanadian Institutes of Health ResearchSanofiOsteoporosis CanadaMcGill UniversityAmgen
Mots-clésFragilityIdentification (biology)Computer scienceMachine learningAlgorithmArtificial intelligenceChemistry

Résumé

récupéré en direct d'OpenAlex

Abstract In this study, we developed ML algorithms to predict fragility fractures, considering the occurrence of fractures at different skeletal sites. We investigated seven ML algorithms (LASSO, Elastic Net, Random Forest, Decision Tree, Neural Network, XGBoost and Logistic Regression) using the data from the Canadian Multicentre Osteoporosis Study (CaMos) with participants aged 50 years or older. We considered 73 baseline features, including age, sex, menopause status, and bone mineral density (BMD), and the outcome was the first incidence of fracture at any of the following sites: hip, spine, pelvis, ribs, shoulder, and forearm, over a 19-year follow-up period. Data were divided into training (70%) and testing (30%) datasets. The ML algorithms were trained on the training dataset and evaluated on the test dataset in terms of the ROC_AUC. SHapley Additive exPlanations (SHAP) analysis was performed to identify the important features that contribute to the prediction of fracture, and to investigate the interaction among these features. In total, 7,753 subjects were included in the study. Approximately 72% were female, and the average age was 67 years. We found that the XGBoost algorithm had a slightly better ROC_AUC (0.70; 95% CI: 0.67, 0.73). From the SHAP analysis, we found that BMD was the most important feature that contributed to the prediction. The other important features include age, previous fracture, osteoporosis and menopausal status. Total hip BMD interacted the most with femoral neck BMD, lumbar spine BMD interacted the most with weight, previous fracture status interacted the most with femoral neck BMD, and age interacted the most with lumbar spine BMD. This study demonstrated that XGBoost was the most effective algorithm for predicting fragility fractures. In addition, we identified important features that contribute to the prediction of fragility fractures. Intervention focusing on these features will help to prevent the incidence of these fractures. Lay summaries We developed machine learning (ML) algorithms to predict fragility fractures, considering the incidence of fractures at different skeletal sites, including the hip, spine, pelvis, ribs, shoulder, or forearm, using 19 years of follow-up data from the Canadian Multicentre Osteoporosis Study (CaMos). We investigated seven ML algorithms and found that XGBoost had slightly better performance compared to other algorithms. We identified important factors that increase the risk of fractures, including BMD, age, and previous fracture. We also demonstrated how the interaction between these factors increases the risk of fractures. The intervention focusing on these factors will help to prevent fragility fractures.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction distillée sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.

score de la tête « metaresearch » (Codex)0,001
score de la tête « metaresearch » (Gemma)0,000
Version: codex-gemma-dda1882f352aStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Simulation ou modélisation · Signal consensuel: Simulation ou modélisation
GenreSignal candidat: Empirique · Signal consensuel: Empirique
Score de désaccord entre enseignants0,519
Score d'incertitude au seuil0,910

Scores Codex et Gemma par catégorie

CatégorieCodexGemma
Métarecherche0,0010,000
Méta-épidémiologie (sens strict)0,0000,000
Méta-épidémiologie (sens large)0,0000,000
Bibliométrie0,0000,000
Études des sciences et des technologies0,0000,000
Communication savante0,0000,000
Science ouverte0,0000,000
Intégrité de la recherche0,0000,001
Charge utile insuffisante (le modèle a refusé de juger)0,0000,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,008
Tête enseignante GPT0,253
Écart entre enseignants0,245 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeSimulation ou modélisation
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations0
Publié2025
Routes d'admission3
Résumé présentoui

Explorer davantage

Même revuemedRxivMême sujetMedical Imaging and AnalysisTravaux en français237 207