Phenotype and genotype comparison of C57BL/6N substrains contributing to the International Mouse Phenotyping Consortium (IMPC)
Bibliographic record
Abstract
The International Mouse Phenotyping Consortium (IMPC) is a 10‐ year program to produce a null mutant mouse line for every gene in the mouse genome, generate comprehensive phenotype data for each, and provide all the resources to the scientific community. The chosen mouse strain is the C57BL/6N, however there are several substrains of this mouse including C57BL/6NCrl (NCrl), C57BL/6NTac (NTac), and C57BL/6NJ (NJ). A C57BL/6Nderived embryonic stem (ES) cell line (JM8) was used to create the gene‐targeted clones. Given the scope of the IMPC project, it is important to catalog phenotype and genotype similarities or differences between the C57BL/6N substrains and ES cells used by the various IMPC centers. Here we describe results from genotype analyses using the Affymetrix Mouse Diversity Genotyping array to compare single‐nucleotide polymorphisms (SNP) between NCrl and NTac mice, and JM8 and C2 ES cell lines. In addition, whole genome sequencing (WGS) was carried out on NCrl, NTac, NJ, and JM8 genomic DNA. Phenotype analyses using the IMPC pipeline was conducted on NCrl and NTac mice at two separate locations; The Toronto Centre for Phenogenomics (TCP) and the Institut Clinique de la Souris (ICS). Substrain differences in certain traits were detected and the number of differences was strongly influenced by the rearing environment of the cohort, with fewer differences between NCrl and NTac cohorts derived from animals born and raised within the same institution.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".