Founder pathogenic variants in colorectal neoplasia susceptibility genes in Ashkenazi Jews undergoing colonoscopy
Bibliographic record
Abstract
BACKGROUND: Colorectal neoplasia is one of the most common tumors affecting Western populations. METHODS: In this study we used a custom amplicon sequencing platform and an in-house bioinformatic pipeline to study constitutional DNA from two different case series of Ashkenazi Jews undergoing colonoscopy (n = 765). The first series all had pathologically confirmed colorectal adenomas and/or carcinoma. The second series consisted of persons who had undergone a colonoscopy within the five years prior to ascertainment, regardless of findings. Ninety-one percent of all patients were asymptomatic at the time of colonoscopy. RESULTS: In the first group (n = 438), we identified 65 founder variants (56 in APC, 2 in GREM1, 3 in MSH2 and 4 in BLM). In the second group (n = 327), the findings were 30, nothing, 1 and 1, respectively, as well as 2 MSH6 variants. CONCLUSIONS: Overall, we found that 10 to 15% of Ashkenazi Jewish persons undergoing colonoscopy harbor variants of interest in colorectal and/or polyposis predisposition. This includes pathogenic variants in MSH6, which is associated with colorectal cancer but not with polyposis. We identified no pathogenic variants in more recently discovered polyposis predisposition genes (POLE, POLD1 or NTHL1), rendering the presence of such founder variants rare.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".