Abstract 2620: Ignoring left truncation in overall survival within real-world genomic-phenomic data leads to inflated survival estimates
Notice bibliographique
Résumé
Abstract Studies linking genomic and phenomic data are subject to selection biases, including delayed entry or immortal time bias. Delayed entry can be problematic for time-to-event analyses, but utilization of appropriate statistical methods to account for delayed entry are underutilized. Delayed entry commonly occurs when genomic sequencing results are obtained after the start time for survival estimation. To evaluate the impact of left truncation on overall survival (OS) estimates, we explored outcomes in patients with de novo stage IV non-small cell lung cancer (NSCLC) and colorectal cancer (CRC) from the AACR GENIE Biopharma Collaborative, who had genomic sequencing within a specified timeframe. We analyzed OS from diagnosis and from start of the most common first-line regimen, carboplatin/pemetrexed for NSCLC (N = 212 patients) and FOLFOX for CRC (N = 369 patients). We compared median OS using standard Kaplan-Meier methods to median OS using left truncation methods to account for delayed entry. All NSCLC and CRC patients underwent genomic sequencing after their diagnosis date. Among NSCLC patients on carboplatin/pemetrexed, 41% and among CRC patients on FOLFOX, 14% had sequencing determined after starting first-line regimen. The survfit function in R package survival was used, and the absolute differences and percent differences in median OS estimates were calculated. Failure to account for delayed entry leads to an overestimation of OS, regardless of cohort and start date. Adjusting survival outcomes using left truncation methods reduces the influence of some aspects of selection bias and results in better estimates of time to event outcomes. Analyses from these cohorts can provide meaningful insights about survival outcomes outside the clinical trial setting and may support trial design and reliable selection of control arms. As such, it is imperative that analytic methods to account for the inflated survival estimates are incorporated. EstimateCRC Stage IV (N = 658)NSCLC Stage IV (N = 722)Unadjusted Median (IQR) Overall Survival from Diagnosis (Years)3.2 (2.9, 3.4)2.3 (2.0, 2.5)Median (IQR) Overall Survival from Diagnosis in Years, Adjusting for Delayed Entry2.1 (1.9, 2.4)1.3 (1.1, 1.6)Difference in Medians (Years)1.11.0% Difference in Medians34%44%EstimateCRC Stage IV (N = 369)NSCLC Stage IV (N = 212)Unadjusted Median (IQR) Overall Survival from Most Common First-Line Regimen (Years)2.9 (2.6, 3.4)1.3 (1.0, 1.6)Median (IQR) Overall Survival from Most Common First-Line Regimen in Years, Adjusting for Delayed Entry2.1 (1.8, 2.5)0.9 (0.7, 1.2)Difference in Medians (Years)0.80.4% Difference in Medians28%31% Citation Format: Samantha Brown, Jessica A. Lavery, Eva M. Lepisto, Caroline McCarthy, Hira Rizvi, Celeste Yu, Kenneth L. Kehl, Shawn M. Sweeney, Julia E. Rudolph, Nikolaus Schultz, Ritika Kundra, Brooke Mastrogiacomo, Phillipe Bedard, Jeremy L. Warner, Gregory J. Riely, Deborah Schrag, Katherine S. Panageas, The AACR Project GENIE Consortium. Ignoring left truncation in overall survival within real-world genomic-phenomic data leads to inflated survival estimates [abstract]. In: Proceedings of the American Association for Cancer Research Annual Meeting 2021; 2021 Apr 10-15 and May 17-21. Philadelphia (PA): AACR; Cancer Res 2021;81(13_Suppl):Abstract nr 2620.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,112 | 0,249 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,001 |
| Méta-épidémiologie (sens large) | 0,002 | 0,003 |
| Bibliométrie | 0,001 | 0,002 |
| Études des sciences et des technologies | 0,001 | 0,002 |
| Communication savante | 0,003 | 0,002 |
| Science ouverte | 0,002 | 0,002 |
| Intégrité de la recherche | 0,001 | 0,003 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,004 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».