Testing for HER2-positive breast cancer: a systematic review and cost-effectiveness analysis
Bibliographic record
Abstract
BACKGROUND: Testing to determine HER2 status has come into focus since the approval of trastuzumab (Herceptin) for the treatment of HER2-positive breast cancer. We compared the cost-effectiveness of various strategies used to test HER2 status, an important first step toward evaluating the overall cost-effectiveness of trastuzumab therapy. METHODS: We performed a systematic review of studies that evaluated concordance between immunohistochemistry and fluorescence in situ hybridization testing to determine HER2 status. We performed a meta-analysis to estimate the distribution of immunohistochemistry scores in each category (0, 1+, 2+, 3+) and the probability of receiving a positive result of fluorescence in situ hybridization (which we assumed to be the "gold-standard" test) for each category. We calculated the accuracy and incremental cost per accurate diagnosis for each testing strategy compared with the base strategy (immunohistochemistry testing, followed by confirmation of 2+ scores by fluorescence in situ hybridization). RESULTS: The median percentage of patients in each category of immunohistochemistry score was: 0, 36.1%; 1+, 35.5%; 2+, 12.0%; and 3+, 16.2%. The median percentage of results of fluorescence in situ hybridization that were positive in each immunohistochemistry category was: 0, 1.6%; 1+, 4.9%; 2+, 29.8%; and 3+, 92.4%. The base strategy was expected to correctly determine the HER2 status of 96% of patients with breast cancer. Confirmation of the HER2 status by fluorescence in situ hybridization in cases that received a score of 3+ reduced the percentage of false-positive results to 0% and increased the percentage of accurately determined HER2 results to 97.6%. Compared with the base strategy, this strategy was associated with a median incremental cost-effectiveness ratio of $6175 per case of accurately determined HER2 status. The strategy of performing fluorescence in situ hybridization testing in all cases of breast cancer was associated with a median incremental cost-effectiveness ratio of $8401 per case of accurately determined HER2 status. INTERPRETATION: The strategy with the lowest cost-effectiveness ratio involved screening all newly diagnosed cases of breast cancer with immunohistochemistry and confirming scores of 2+ or 3+ with fluorescence in situ hybridization testing.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.013 | 0.013 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.005 | 0.001 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.001 | 0.002 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".