Chromogenic in situ hybridisation for the assessment of HER2 status in breast cancer: an international validation ring study
Bibliographic record
Abstract
INTRODUCTION: Before any new methodology can be introduced into the routine diagnostic setting it must be technically validated against the established standards. To this end, a ring study involving five international pathology laboratories was initiated to validate chromogenic in situ hybridisation (CISH) against fluorescence in situ hybridisation (FISH) and immunohistochemistry (IHC) as a test for assessing human epidermal growth factor receptor 2 (HER2) status in breast cancer. METHODS: Each laboratory performed CISH, FISH and IHC on its own samples. Unstained sections from each case were also sent to another participating laboratory for blinded retesting by CISH ('outside CISH'). RESULTS: A total of 211 invasive breast carcinoma cases were tested. In 76 cases with high amplification (HER2/CEP17 ratio >4.0) by FISH, 73 cases (96%) scored positive (scores >or= 6) by 'outside CISH'. For FISH-negative cases (HER2/CEP17 ratio <2.0), 94 of 100 cases (94%) had CISH scores indicating no amplification (score <or= 5), and only three cases were positive by CISH; in the three remaining cases, no CISH result could be obtained. For cases with low-level amplification using FISH (HER2/CEP17 ratio 2.0-4.0), 20 of 35 had CISH scores indicating gene amplification. Inter-laboratory concordance was also very high: 95% for normal HER2 copy number (1-5 copies); and 92% for cases with HER2 copy numbers >or= 6. CISH intra-laboratory concordance with IHC was 92% for IHC-negative cases (IHC 0/1+) and 91% for IHC 3+ cases. Among IHC 2+ cases, CISH was 100% concordant with samples showing high amplification by FISH, and 94% concordant with FISH-negative samples. CONCLUSION: These results show that CISH inter- and intra-laboratory concordance to FISH and IHC is very high, even in equivocal IHC 2+ cases. Therefore, we conclude that CISH is a methodology that is a viable alternative to FISH in the HER2 testing algorithm.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.007 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".