Additional file 1 of Homozygote CRIM1 variant is associated with thiopurine-induced neutropenia in leukemic patients with both wildtype NUDT15 and TPMT
Bibliographic record
Abstract
Additional file 1: Table S1. Evaluation of 12 candidate variants from the discovery cohort (N = 188) by using the replication cohort (N = 52) for both NUDT15 and TPMT wild-type subjects. Table S2. Evaluation of frequency distributions of CRIM1 rs3821169 genotypes across different cutoffs of the last-cycle 6-mercaptopurine dose intensity percentage tolerated by pediatric acute lymphoblastic leukemia subjects. Figure S1. Improvement of prediction accuracy of GVBNUDT15,TPMT for thiopurine toxicity after controlling for homozygote carriers of CRIM1 rs3821169. Figure S2. Prediction accuracies of GVBNUDT15,CRIM1 and GVBTPMT,CRIM1 for thiopurine toxicity in pediatric ALL subjects. Figure S3. Prediction accuracies of GVBNUDT15,TPMT,CRIM1 for thiopurine toxicity in pediatric ALL subjects. Figure S4. Prediction accuracy of GVBCRIM1 for thiopurine toxicity in pediatric ALL subjects. Figure S5. Prediction accuracy of GVBNUDT15 for thiopurine toxicity in pediatric ALL subjects. Figure S6. Prediction accuracy of GVBTPMT for thiopurine toxicity in pediatric ALL subjects. Figure S7. Youden’s index to find the optimal thresholds for GVBNUDT15,TPMT and GVBNUDT15,TPMT,CRIM1. Figure S8. Comparison of CRIM1 mRNA expression levels of rs3821169 carriers and noncarriers in hematopoietic and lymphoid tissue. Figure S9. Results of Sanger sequencing for the two NUDT15 variants identified via whole exome sequencing.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.028 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.002 | 0.003 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.002 | 0.002 |
| Open science | 0.002 | 0.001 |
| Research integrity | 0.002 | 0.001 |
| Insufficient payload (model declined to judge) | 0.794 | 0.066 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".