ESMO Magnitude of Clinical Benefit Scale (MCBS): An evaluation of systemic treatment trials for soft tissue sarcomas (STS).
Bibliographic record
Abstract
11553 Background: Patients with STS have poor prognosis in the metastatic setting. Although some treatment options are associated with improved outcomes, such as progression-free (PFS) or overall survival (OS), the overall magnitude of clinical benefit can be unclear. The ESMO MCBS is a validated and reproducible tool developed to quantify the clinical benefit of treatments evaluated in trials ( www.esmo.org/guidelines/esmo-mcbs ). Herein, we report the application of ESMO MCBS to systemic treatment trials involving metastatic STS patients. Methods: A systematic search of Medline, Embase and Cochrane databases for adult phase II and III trials in advanced STS (01/1998 to 12/2020) was carried out. Gastrointestinal stromal tumor trials were excluded. Outcomes, including but not limited to OS, PFS, objective response rate (ORR), toxicity and quality of life (QoL) data were extracted and analyzed. Studies with outcomes that met the criteria for ESMO MCBS v1.1 were evaluated to generate a score of 1 to 5 (score of ≥ 4: substantial benefit). MCBS scoring of each study was performed by at least 2 co-authors for consensus. Results: Among 3454 abstracts screened, a total of 140 Phase II and 28 phase IIII trials were identified. A total of 41 studies fulfilled the criteria for ESMO MCBS scoring. These include 5 phase III studies, as well as 9 randomized and 27 single-arm phase II trials. Fifteen studies involved specific histology, while remaining 26 studies were of all STS subtypes. Chemotherapy, alone or in combination was evaluated in 29 trials, while molecular-targeted agents (MTA) and immune checkpoint inhibitors (IO) were evaluated in 11 and 3 studies, respectively (Table). The median MCBS score was 2 (range 1-4), regardless of drug class or combination. Only 3 studies, all randomized in design, had a MCBS score of 4. All three trials were in the 2nd line setting or beyond, where there is no standard control treatment. None of the trials, irrespective of drug class had a score of 5 and no study showed evidence of significant improvement in QoL. The observed MCBS scores were low, partly because the trials evaluated mainly comprise single-arm studies without QoL assessments, restricting to a maximum MCBS score of 3. Conclusions: Most systemic therapy trials in advanced STS did not confer substantial clinical benefit when evaluated using MCBS. Although randomized phase 3 trials remain the gold standard of treatment evaluation, clinical benefit evaluation of STS trials using tools such as MCBS may be useful. Incorporation of QoL evaluation, even in single-arm studies should be prioritized in metastatic STS trials.[Table: see text]
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.047 | 0.106 |
| Meta-epidemiology (narrow) | 0.002 | 0.001 |
| Meta-epidemiology (broad) | 0.009 | 0.013 |
| Bibliometrics | 0.014 | 0.011 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.003 | 0.003 |
| Open science | 0.002 | 0.003 |
| Research integrity | 0.002 | 0.001 |
| Insufficient payload (model declined to judge) | 0.006 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".