Prospective evaluation of a breast-cancer risk model integrating classical risk factors and polygenic risk in 15 cohorts from six countries
Bibliographic record
Abstract
BACKGROUND: Rigorous evaluation of the calibration and discrimination of breast-cancer risk-prediction models in prospective cohorts is critical for applications under clinical guidelines. We comprehensively evaluated an integrated model incorporating classical risk factors and a 313-variant polygenic risk score (PRS) to predict breast-cancer risk. METHODS: Fifteen prospective cohorts from six countries with 239 340 women (7646 incident breast-cancer cases) of European ancestry aged 19-75 years were included. Calibration of 5-year risk was assessed by comparing expected and observed proportions of cases overall and within risk categories. Risk stratification for women of European ancestry aged 50-70 years in those countries was evaluated by the proportion of women and future cases crossing clinically relevant risk thresholds. RESULTS: Among women <50 years old, the median (range) expected-to-observed ratio for the integrated model across 15 cohorts was 0.9 (0.7-1.0) overall and 0.9 (0.7-1.4) at the highest-risk decile; among women ≥50 years old, these were 1.0 (0.7-1.3) and 1.2 (0.7-1.6), respectively. The proportion of women identified above a 3% 5-year risk threshold (used for recommending risk-reducing medications in the USA) ranged from 7.0% in Germany (∼841 000 of 12 million) to 17.7% in the USA (∼5.3 of 30 million). At this threshold, 14.7% of US women were reclassified by adding the PRS to classical risk factors, with identification of 12.2% of additional future cases. CONCLUSION: Integrating a 313-variant PRS with classical risk factors can improve the identification of European-ancestry women at elevated risk who could benefit from targeted risk-reducing strategies under current clinical guidelines.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.007 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".