MétaCan
Menu
Back to cohort

Measuring the long-term “tail of curve” survival benefits in oncology trials: A comparison of the ASCO Value Framework and the ESMO Magnitude of Clinical Benefit Scale.

2019· article· en· W2947154035 on OpenAlexaff
Louis Everest, Monica Shah, Kelvin Chan

Bibliographic record

VenueJournal of Clinical Oncology · 2019
Typearticle
Languageen
FieldEconomics, Econometrics and Finance
TopicEconomic and Financial Impacts of Cancer
Canadian institutionsHealth Sciences CentreSunnybrook Health Science Centre
Fundersnot available
KeywordsMedicineOncologyInternal medicineClinical trialRandomized controlled trialImmunotherapyClinical OncologyCancer

Abstract

fetched live from OpenAlex

2585 Background: Recently, anti-cancer agents have generated excitement due to their capacity to preserve long-term survival in some patients, represented by a “tail of the survival curve”. However, as traditional measures of clinical benefit may not accurately capture long-term survival, amendments to various valuation frameworks have been proposed to capture this benefit. The purpose of this study was to determine how frequently immune checkpoint inhibitor vs. non-immune checkpoint inhibitor anti-cancer agents, displayed trends of long-term survival, as defined by the American Society of Clinical Oncology Value Framework (ASCO-VF) and European Society of Medical Oncology Magnitude of Clinical Benefit Scale (ESMO-MCBS), as well as to analyze the degree of agreement between ASCO and ESMO frameworks. Methods: Anti-cancer agents from phase II or III randomized controlled trials (RCTs) cited for clinical efficacy evidence in drug approval by the Food and Drug Administration (FDA) between January 2011 and March 2018 were identified. Data required for ASCO-VF and ESMO-MCBS were extracted. Difference in how often long-term survival bonuses were awarded were calculated in all RCTs, as well as immune checkpoint inhibitor and non-immune checkpoint inhibitor RCTs individually. Cohen’s Kappa statistic was calculated to evaluate agreement between ASCO-VF and ESMO-MCBS. Results: 100 RCTs were analyzed. RCTs were awarded ASCO-VF version 2 (v2) “tail of the curve” bonuses more often than ESMO-MCBS version 1.1 (v1.1) “immunotherapy-triggered” long-term plateau adjustments (45% vs. 2.6%). Comparing to non-immune checkpoint inhibitor RCTs, immune checkpoint inhibitor RCTs were not more likely to receive ASCO-VF v2 bonuses/ESMO-MCBS v1.1 adjustments (p = 0.32/ p = 0.40). Long-term survival agreement between the two frameworks was poor (kappa: 0.01; p = 0.50). Conclusions: The ASCO-VF v2 and ESMO-MCBS v1.1 may require additional refinement in order to accurately capture the benefit of long-term survival or immune checkpoint inhibitor and non-immune checkpoint inhibitor agents may not preserve substantially different long-term survival populations.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.387
metaresearch head score (Gemma)0.651
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch
Consensus categoriesMetaresearch
DomainCandidate signal: Evaluation · Consensus signal: none
Study designCandidate signal: Theoretical or conceptual · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.613
Threshold uncertainty score0.756

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.3870.651
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0040.007
Bibliometrics0.0170.018
Science and technology studies0.0010.003
Scholarly communication0.0050.006
Open science0.0020.007
Research integrity0.0020.003
Insufficient payload (model declined to judge)0.0030.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.308
GPT teacher head0.458
Teacher spread0.149 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; the direct Gemma label and the distilled Codex classifier agree on what is shown here.

Study designTheoretical or conceptual
DomainEvaluation
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations1
Published2019
Admission routes1
Has abstractyes

Explore more

Same venueJournal of Clinical OncologySame topicEconomic and Financial Impacts of CancerFrench-language works237,207