Analysis of the Cat Eye Syndrome Critical Region in Humans and the Region of Conserved Synteny in Mice: A Search for Candidate Genes at or near the Human Chromosome 22 Pericentromere
Bibliographic record
Abstract
We have sequenced a 1.1-Mb region of human chromosome 22q containing the dosage-sensitive gene(s) responsible for cat eye syndrome (CES) as well as the 450-kb homologous region on mouse chromosome 6. Fourteen putative genes were identified within or adjacent to the human CES critical region (CESCR), including three known genes (IL-17R, ATP6E, and BID) and nine novel genes, based on EST identity. Two putative genes (CECR3 and CECR9) were identified, in the absence of EST hits, by comparing segments of human and mouse genomic sequence around two solitary amplified exons, thus showing the utility of comparative genomic sequence analysis in identifying transcripts. Of the 14 genes, 10 were confirmed to be present in the mouse genomic sequence in the same order and orientation as in human. Absent from the mouse region of conserved synteny are CECR1, a promising CES candidate gene from the center of the contig, neighboring CECR4, and CECR7 and CECR8, which are located in the gene-poor proximal 400 kb of the contig. This latter proximal region, located approximately 1 Mb from the centromere, shows abundant duplicated gene fragments typical of pericentromeric DNA. The margin of this region also delineates the boundary of conserved synteny between the CESCR and mouse chromosome 6. Because the proximal CESCR appears abundant in duplicated segments and, therefore, is likely to be gene poor, we consider the putative genes identified in the distal CESCR to represent the majority of candidate genes for involvement in CES.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.002 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".