A pilot study of machine-learning based automated planning for primary brain tumours
Bibliographic record
Abstract
PURPOSE: High-quality radiotherapy (RT) planning for children and young adults with primary brain tumours is essential to minimize the risk of late treatment effects. The feasibility of using automated machine-learning (ML) to aid RT planning in this population has not previously been studied. METHODS AND MATERIALS: We developed a ML model that identifies learned relationships between image features and expected dose in a training set of 95 patients with a primary brain tumour treated with focal radiotherapy to a dose of 54 Gy in 30 fractions. This ML method was then used to create predicted dose distributions for 15 previously-treated brain tumour patients across two institutions, as a testing set. Dosimetry to target volumes and organs-at-risk (OARs) were compared between the clinically-delivered (human-generated) plans versus the ML plans. RESULTS: The ML method was able to create deliverable plans in all 15 patients in the testing set. All ML plans were generated within 30 min of initiating planning. Planning target volume coverage with 95% of the prescription dose was attained in all plans. OAR doses were similar across most structures evaluated; mean doses to brain and left temporal lobe were lower in ML plans than manual plans (mean difference to left temporal, - 2.3 Gy, p = 0.006; mean differences to brain, - 1.3 Gy, p = 0.017), whereas mean doses to right cochlea and lenses were higher in ML plans (+ 1.6-2.2 Gy, p < 0.05 for each). CONCLUSIONS: Use of an automated ML method to aid RT planning for children and young adults with primary brain tumours is dosimetrically feasible and can be successfully used to create high-quality 54 Gy RT plans. Further evaluation after clinical implementation is planned.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".