Radiomic-based approaches in the multi-metastatic setting: a quantitative review
Bibliographic record
Abstract
BACKGROUND: Radiomics traditionally focuses on analyzing a single lesion within a patient to extract tumor characteristics, yet this process may overlook inter-lesion heterogeneity, particularly in the multi-metastatic setting. There is currently no established method for combining radiomic features in such settings, leading to diverse approaches with varying strengths and limitations. Our quantitative review aims to illuminate these methodologies, assess their replicability, and guide future research toward establishing best practices, offering insights into the challenges of multi-lesion radiomic analysis across diverse datasets. METHODS: We conducted a comprehensive literature search to identify methods for integrating data from multiple lesions in radiomic analyses. We replicated these methods using either the author's code or by reconstructing them based on the information provided in the papers. Subsequently, we applied these identified methods to three distinct datasets, each depicting a different metastatic scenario. RESULTS: We compared ten mathematical methods for combining radiomic features across three distinct datasets, encompassing 16,894 lesions in 3,930 patients. Performance was evaluated using the Cox proportional hazards model and benchmarked against univariable analysis of total tumor volume. Results varied by dataset and lesion burden, with no single method consistently outperforming others. In colorectal liver metastases (TCIA-CRLM, 494 lesions in 197 patients), averaging methods showed the highest median performance. In soft tissue sarcoma (TH CR-406/SARC021, 1255 lesions in 545 patients), concatenating radiomic features from multiple lesions exhibited the best performance. In head and neck cancers (TCIA-RADCURE, 15,145 lesions in 3188 patients), total tumor volume remained a strong predictor. These findings highlight dataset-specific influences, including tumor type and lesion burden, on the effectiveness of radiomic feature aggregation methods. CONCLUSIONS: Radiomic features can be effectively selected or combined to estimate patient-level outcomes in multi-metastatic patients, though the approach varies by metastatic setting. Our study fills a critical gap in radiomics research by examining the challenges of radiomic-based analysis in this setting. Through a comprehensive review and rigorous testing of different methods across diverse datasets representing unique metastatic scenarios, we provide valuable insights into effective radiomic analysis strategies.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".