Case-control study of mammographic density and breast cancer risk using processed digital mammograms
Bibliographic record
Abstract
BACKGROUND: Full-field digital mammography (FFDM) has largely replaced film-screen mammography in the US. Breast density assessed from film mammograms is strongly associated with breast cancer risk, but data are limited for processed FFDM images used for clinical care. METHODS: We conducted a case-control study nested among non-Hispanic white female participants of the Research Program in Genes, Environment and Health of Kaiser Permanente Northern California who were aged 40 to 74 years and had screening mammograms acquired on Hologic FFDM machines. Cases (n = 297) were women with a first invasive breast cancer diagnosed after a screening FFDM. For each case, up to five controls (n = 1149) were selected, matched on age and year of FFDM and image batch number, and who were still under follow-up and without a history of breast cancer at the age of diagnosis of the matched case. Percent density (PD) and dense area (DA) were assessed by a radiological technologist using Cumulus. Conditional logistic regression was used to estimate odds ratios (ORs) for breast cancer associated with PD and DA, modeled continuously in standard deviation (SD) increments and categorically in quintiles, after adjusting for body mass index, parity, first-degree family history of breast cancer, breast area, and menopausal hormone use. RESULTS: Median intra-reader reproducibility was high with a Pearson's r of 0.956 (range 0.902 to 0.983) for replicate PD measurements across 23 image batches. The overall mean was 20.02 (SD, 14.61) for PD and 27.63 cm(2) (18.22 cm(2)) for DA. The adjusted ORs for breast cancer associated with each SD increment were 1.70 (95 % confidence interval, 1.41-2.04) for PD, and 1.54 (1.34-1.77) for DA. The adjusted ORs for each quintile were: 1.00 (ref.), 1.49 (0.91-2.45), 2.57 (1.54-4.30), 3.22 (1.91-5.43), 4.88 (2.78-8.55) for PD, and 1.00 (ref.), 1.43 (0.85-2.40), 2.53 (1.53-4.19), 2.85 (1.73-4.69), 3.48 (2.14-5.65) for DA. CONCLUSIONS: PD and DA measured using Cumulus on processed FFDM images are positively associated with breast cancer risk, with similar magnitudes of association as previously reported for film-screen mammograms. Processed digital mammograms acquired for routine clinical care in a general practice setting are suitable for breast density and cancer research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".