From Gleason to International Society of Urological Pathology (ISUP) grading of prostate cancer
Bibliographic record
Abstract
Gleason grading of prostate cancer has gained worldwide acceptance since its introduction 50 years ago. This system has fulfilled the role of a powerful prognostic indicator for many years and this has influenced treatment. There have been numerous changes to the management and diagnosis of prostate cancer since 1966, including prostate-specific antigen screening, resulting in the early detection of prostate cancer, This has resulted in the evolution of Gleason grading with the informal adoption of a number of alterations. Significant changes to Gleason grading were made in 2005 through a consensus conference convened by the International Society of Urological Pathology (ISUP). In more recent times, the necessity for further changes to prostate cancer grading has been apparent and a follow-up ISUP consensus conference was held in 2014. Changes resulting from this conference included the classifying of all cribriform cancer and glomeruloid patterns as Gleason grade 4, the grading of mucinous adenocarcinoma based on underlying architecture rather than uniformly considering these tumors as pattern 4, and the introduction of a Gleason score (GS)-based 5 grade system, which incorporated the 2014 modifications to the Gleason grading system. Designated ISUP grade, this system consists of five grades: grade 1 (GS ≤3 + 3), grade 2 (GS 3 + 4), grade 3 (GS 4 + 3), grade 4 (GS 4 + 4, 3 + 5, 5 + 3) and grade 5 (GS 9-10). With further advances recently reported in the literature, it is apparent that amendments to the current system are likely to be necessary in the future.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.003 | 0.001 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".