Evaluation of 16 loci to examine the cross‐species utility of single nucleotide polymorphism arrays
Bibliographic record
Abstract
Large collections of single nucleotide polymorphisms (SNPs) have recently been identified from a number of livestock genomes. This raises the possibility that SNP arrays might be useful for analysis in related species for which few genetic markers are currently available. To address the likely success of such an approach, the aim of this study was to examine the threshold number and position of flanking mutations which act to prevent genotype calls being produced. Sequence diversity was measured across 16 loci containing SNPs known either to work successfully between species or fail between species. In pairwise comparisons between domestic and wild sheep, sequence divergence surrounding working SNP assays was significantly lower than that surrounding non-functional assays. In addition, the location of flanking mismatches tended to be closer to the target SNP in loci that failed to generate genotype calls across species. The magnitude of sequence divergence observed for both working and non-functional assays was compared with the divergence separating domestic sheep from European Mouflon, African Barbary, goat and cattle. The results suggest that the utility of SNP arrays for analysis of shared polymorphism will be restricted to closely related pairs of species. Analysis across more divergent species will, however, be successful for other objectives, such as the identification of the ancestral state of SNPs.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".