MétaCan
Menu
← Retour à la cohorte
Enregistrement W4405041653 · doi:10.1182/blood-2024-210900

Predictive Modeling of Clonal Hematopoiesis across Diverse Cohorts

2024· article· en· W4405041653 sur OpenAlexaff
Armel L. Batchi‐Bouyou, Duc Tran, Michael A. Raddatz, Teng Gao, Caitlyn Vlasschaert, Brian Sharber, Jie Liu, Giulia E.M. Petrone, Irenaeus C.C. Chan, J. Scott Beeler, Griffen Mustion, Sean M. Devlin, Alexander G. Bick, Kelly L. Bolton

Notice bibliographique

RevueBlood · 2024
Typearticle
Langueen
DomaineMedicine
ThématiqueMyeloproliferative Neoplasms: Diagnosis and Treatment
Établissements canadiensQueen's University
Organismes subventionnairesnon disponible
Mots-clésBiologyHaematopoiesisImmunologyComputational biologyMedicineGeneticsStem cell

Résumé

récupéré en direct d'OpenAlex

Background: Clonal hematopoiesis (CH) is associated an increased risk of hematologic malignancy and numerous other adverse events. While there are no established therapeutic approaches to treat CH, clinical trials are underway, many of which use targeted therapeutic approaches for specific CH genes or genetic pathways. Since CH screening is generally not part of routine clinical testing, the identification of CH-positive individuals is a challenge to recruitment. While age is the most important risk factor for CH, CH is also known to be associated with other demographic and clinical risk factors. Blood count and blood parameters are also known to be influenced by CH with gene-specific patterns. Here we sought to understand whether commonly available clinical predictors and blood counts could be used to develop a gene-specific CH risk prediction tool with the goal of facilitating CH research screening strategies. Methods: We constructed a prediction model for CH across three cohorts of participants using LASSO regression. Clinical predictors included clinical/demographic characteristics (smoking history, gender, race) and blood count parameters. The UK Biobank served as the development cohort, consisting of 452,547 participants. We used two separate cohorts for validation, The All of Us (AoU) Research database, consisting of 143,850 participants, and MSK-IMPACT cohort, including 8,150 patients with non-hematologic cancers. We compared models with age alone to models including blood count and other clinical/demographic parameters. The predictive performance was determined based on 2 criteria: discrimination by calculating the area under the curve (AUC) receiver operating characteristic (ROC) and calibration by calculating the calibration slope (slope of 1 indicates perfect calibration) and the intercept. Results: A total of 604,547 participants were included in the study. We observed strong associations between clinical features and gene-specific CH including platelet count with DNMT3A and JAK2, neutrophil count and IDH1/2 mutations, and a strong association between spliceosome CH and age. Overall our model showed excellent discrimination (AUC>0.8) for risk JAK2, ASXL1, PPM1D, SF3B1, SRF2, U2AF1 and modest discrimination (AUC>0.7) for DNMT3A, IDH1/2, TP53 and TET2. Compared to a model with age alone, the addition of blood count and clinical parameters improved the model's performance most notably for JAK2 (AUC = 0.72 vs 0.82) and IDH1/2 (AUC = 0.75 vs 0.78). The calibration slopes for gene-specific models ranged from 0.35-1.65 and were highest for JAK2 (slope=0.9; intercept=0.02 ) and TP53 (slope=0.89; intercept=-0.02) . To better determine how our risk prediction model could be used to inform CH screening strategies, we determined the number of patients that would be required to screen using our CH risk prediction model and the number needed to sequence to identify 100 CH positive individuals across 10 CH genes. Application of our risk prediction model to identify individuals at high risk of CH for screening reduced the number of samples needed to sequence by 4-19 fold. Conclusion: We developed and validated a model for gene-specific CH prediction using blood count parameters and demographic factors with strong discriminative performance. These findings highlight the potential of commonly available clinical data to improve CH prediction, aiding in efficient identification of individuals with CH to facilitate clinical trial design.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,010
score de la tête « metaresearch » (Gemma)0,013
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Simulation ou modélisation · Signal consensuel: Simulation ou modélisation
GenreSignal candidat: Empirique · Signal consensuel: Empirique
Score de désaccord entre enseignants0,010
Score d'incertitude au seuil0,052

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,0100,013
Méta-épidémiologie (sens strict)0,0010,000
Méta-épidémiologie (sens large)0,0010,001
Bibliométrie0,0010,001
Études des sciences et des technologies0,0010,001
Communication savante0,0010,001
Science ouverte0,0010,001
Intégrité de la recherche0,0010,002
Charge utile insuffisante (le modèle a refusé de juger)0,0020,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,022
Tête enseignante GPT0,298
Écart entre enseignants0,276 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeSimulation ou modélisation
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations0
Publié2024
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueBlood→Même sujetMyeloproliferative Neoplasms: Diagnosis and Treatment→Travaux en français237 207→