MétaCan
Menu
Back to cohort

Evaluating the value of checkpoint inhibitor therapy using the ASCO and ESMO frameworks.

2019· article· en· W2960915348 on OpenAlexaff
Sophie Feng, Yanshuo Cao, Eitan Amir, Eric Xueyu Chen

Bibliographic record

VenueJournal of Clinical Oncology · 2019
Typearticle
Languageen
FieldEconomics, Econometrics and Finance
TopicEconomic and Financial Impacts of Cancer
Canadian institutionsUniversity Health NetworkPrincess Margaret Cancer Centre
Fundersnot available
KeywordsMedicineClinical OncologyInternal medicineRandomized controlled trialHazard ratioOncologyClinical trialLung cancerCancerConfidence interval

Abstract

fetched live from OpenAlex

17 Background: The advent of checkpoint inhibitor therapy (CIT) has dramatically changed the oncology landscape, but is associated with significant costs. A positive randomized clinical trial (RCT) may not translate to meaningful outcomes for patients. The American Society of Clinical Oncology (ASCO) and European Society of Medical Oncology (ESMO) have developed frameworks to quantify the value of cancer treatment. We applied these frameworks to RCTs involving CIT in order to explore the relationship between trial outcomes and magnitude of clinical benefit. Methods: A literature search was conducted to identify CIT RCTs. Data extracted included study characteristics, pre-specified estimated hazard ratios (eHR) and observed HR (oHR). ASCO Value Framework version 2016 and ESMO Magnitude of Clinical Benefit (MCB) scale v1.1 were applied to each publication by 2 authors independently. Results: 30 RCTs (3 adjuvant, and 27 advanced setting) using CIT were identified between January 2010- October 2018. The majority of trials were in lung cancer (37%) and melanoma (36%). The eHR was 0.71±0.06 (range: 0.55-0.78), and oHR was 0.76±0.15 (range: 0.49 – 1.11). 54% RCTs did not achieve eHR, with a difference of 0.16±0.12 (range: 0.01 – 0.41). ASCO framework scores ranged -24.0 to 71.3, far below the maximum potential score of 180. 18 RCTs formed the basis for FDA approvals. The mean ASCO framework score was 45.6 ± 16.6 (range: 14.4 - 71.3) for FDA approved indications, and 14.0 ± 18.4 (range: -24 – 49) for non-FDA approvals (p < 0.001). All FDA approvals scored grade 4 or 5 on the ESMO MCB scale, indicating a meaningful clinical benefit. Many non-FDA approved RCTs did not receive an MCB grade as they were negative trials. There was no difference in the ASCO framework scores between MCB grade 4 and 5 RCTs. 3 adjuvant RCTs had an ASCO framework score ranging from 20.5 to 38.7 despite an MCB Grade A. Conclusions: Many trials did not meet the pre-specified eHR. FDA approval had statistically significantly higher NHB scores than non-FDA approvals, and they were deemed to have a meaningful clinical benefit according to the ESMO MCB scale. ASCO framework scores may require re-calibration since the highest score achieved in CIT RCTs was only 40% of the maximum score.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.195
metaresearch head score (Gemma)0.358
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.195
Threshold uncertainty score0.993

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.1950.358
Meta-epidemiology (narrow)0.0030.001
Meta-epidemiology (broad)0.0060.018
Bibliometrics0.0210.012
Science and technology studies0.0010.003
Scholarly communication0.0070.004
Open science0.0030.008
Research integrity0.0030.003
Insufficient payload (model declined to judge)0.0050.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.255
GPT teacher head0.473
Teacher spread0.218 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

Study designNot applicable
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2019
Admission routes1
Has abstractyes

Explore more

Same venueJournal of Clinical OncologySame topicEconomic and Financial Impacts of CancerFrench-language works237,207