Race, Breast Cancer Subtypes, and Survival in the Carolina Breast Cancer Study
Bibliographic record
Abstract
CONTEXT: Gene expression analysis has identified several breast cancer subtypes, including basal-like, human epidermal growth factor receptor-2 positive/estrogen receptor negative (HER2+/ER-), luminal A, and luminal B. OBJECTIVES: To determine population-based distributions and clinical associations for breast cancer subtypes. DESIGN, SETTING, AND PARTICIPANTS: Immunohistochemical surrogates for each subtype were applied to 496 incident cases of invasive breast cancer from the Carolina Breast Cancer Study (ascertained between May 1993 and December 1996), a population-based, case-control study that oversampled premenopausal and African American women. Subtype definitions were as follows: luminal A (ER+ and/or progesterone receptor positive [PR+], HER2-), luminal B (ER+ and/or PR+, HER2+), basal-like (ER-, PR-, HER2-, cytokeratin 5/6 positive, and/or HER1+), HER2+/ER- (ER-, PR-, and HER2+), and unclassified (negative for all 5 markers). MAIN OUTCOME MEASURES: We examined the prevalence of breast cancer subtypes within racial and menopausal subsets and determined their associations with tumor size, axillary nodal status, mitotic index, nuclear pleomorphism, combined grade, p53 mutation status, and breast cancer-specific survival. RESULTS: The basal-like breast cancer subtype was more prevalent among premenopausal African American women (39%) compared with postmenopausal African American women (14%) and non-African American women (16%) of any age (P<.001), whereas the luminal A subtype was less prevalent (36% vs 59% and 54%, respectively). The HER2+/ER- subtype did not vary with race or menopausal status (6%-9%). Compared with luminal A, basal-like tumors had more TP53 mutations (44% vs 15%, P<.001), higher mitotic index (odds ratio [OR], 11.0; 95% confidence interval [CI], 5.6-21.7), more marked nuclear pleomorphism (OR, 9.7; 95% CI, 5.3-18.0), and higher combined grade (OR, 8.3; 95% CI, 4.4-15.6). Breast cancer-specific survival differed by subtype (P<.001), with shortest survival among HER2+/ER- and basal-like subtypes. CONCLUSIONS: Basal-like breast tumors occurred at a higher prevalence among premenopausal African American patients compared with postmenopausal African American and non-African American patients in this population-based study. A higher prevalence of basal-like breast tumors and a lower prevalence of luminal A tumors could contribute to the poor prognosis of young African American women with breast cancer.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".