Comprehensive epithelial tubo-ovarian cancer risk prediction model incorporating genetic and epidemiological risk factors
Bibliographic record
Abstract
Background Epithelial tubo-ovarian cancer (EOC) has high mortality partly due to late diagnosis. Prevention is available but may be associated with adverse effects. A multifactorial risk model based on known genetic and epidemiological risk factors (RFs) for EOC can help identify women at higher risk who could benefit from targeted screening and prevention. Methods We developed a multifactorial EOC risk model for women of European ancestry incorporating the effects of pathogenic variants (PVs) in BRCA1 , BRCA2 , RAD51C , RAD51D and BRIP1 , a Polygenic Risk Score (PRS) of arbitrary size, the effects of RFs and explicit family history (FH) using a synthetic model approach. The PRS, PV and RFs were assumed to act multiplicatively. Results Based on a currently available PRS for EOC that explains 5% of the EOC polygenic variance, the estimated lifetime risks under the multifactorial model in the general population vary from 0.5% to 4.6% for the first to 99th percentiles of the EOC risk distribution. The corresponding range for women with an affected first-degree relative is 1.9%–10.3%. Based on the combined risk distribution, 33% of RAD51D PV carriers are expected to have a lifetime EOC risk of less than 10%. RFs provided the widest distribution, followed by the PRS. In an independent partial model validation, absolute and relative 5-year risks were well calibrated in quintiles of predicted risk. Conclusion This multifactorial risk model can facilitate stratification, in particular among women with FH of cancer and/or moderate-risk and high-risk PVs. The model is available via the CanRisk Tool ( www.canrisk.org ).
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".