Organ at risk delineation for radiation therapy clinical trials: Global Harmonization Group consensus guidelines: GHG OAR consensus contouring guidance
Bibliographic record
Abstract
Background and purpose: The Global Quality Assurance of Radiation Therapy Clinical Trials Harmonization Group (GHG) is a collaborative group of Radiation Therapy Quality Assurance (RTQA) Groups harmonizing and improving RTQA for multi-institutional clinical trials. The objective of the GHG OAR Working Group was to unify OAR contouring guidance across RTQA groups by compiling a single reference list of OARs in line with AAPM TG 263 and ASTRO, together with peer-reviewed, anatomically defined contouring guidance for integration into clinical trial protocols independent of the radiation therapy delivery technique. Materials and methods: The GHG OAR Working Group comprised of 22 multi-professional members from 6 international RTQA Groups and affiliated organizations conducted the work in 3 stages: (1) Clinical trial documentation review and identification of structures of interest (2) Review of existing contouring guidance and survey of proposed OAR contouring guidance (3) Review of survey feedback with recommendations for contouring guidance with standardized OAR nomenclature. Results: 157 clinical trials were examined; 222 OAR structures were identified. Duplicates, non-anatomical, non-specific, structures with more specific alternative nomenclature, and structures identified by one RTQA group were excluded leaving 58 structures of interest. 6 OAR descriptions were accepted with no amendments, 41 required minor amendments, 6 major amendments, 20 developed as a result of feedback, and 5 structures excluded in response to feedback. The final GHG consensus guidance includes 73 OARs with peer-reviewed descriptions (Appendix A). Conclusion: We provide OAR descriptions with standardized nomenclature for use in clinical trials. A more uniform dataset supports the delivery of clinically relevant and valid conclusions from clinical trials.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".