Smelt was the likely beneficiary of an antifreeze gene laterally transferred between fishes
Bibliographic record
Abstract
BACKGROUND: Type II antifreeze protein (AFP) from the rainbow smelt, Osmerus mordax, is a calcium-dependent C-type lectin homolog, similar to the AFPs from herring and sea raven. While C-type lectins are ubiquitous, type II AFPs are only found in a few species in three widely separated branches of teleost fishes. Furthermore, several other non-homologous AFPs are found in intervening species. We have previously postulated that this sporadic distribution has resulted from lateral gene transfer. The alternative hypothesis, that the AFP evolved from a lectin present in a shared ancestor and that this gene was lost in most species, is not favored because both the exon and intron sequences are highly conserved. RESULTS: Here we have sequenced and annotated a 160 kb smelt BAC clone containing a centrally-located AFP gene along with 14 other genes. Quantitative PCR indicates that there is but a single copy of this gene within the smelt genome, which is atypical for fish AFP genes. The corresponding syntenic region has been identified and searched in a number of other species and found to be devoid of lectin or AFP sequences. Unlike the introns of the AFP gene, the intronic sequences of the flanking genes are not conserved between species. As well, the rate and pattern of mutation in the AFP gene are radically different from those seen in other smelt and herring genes. CONCLUSIONS: These results provide stand-alone support for an example of lateral gene transfer between vertebrate species. They should further inform the debate about genetically modified organisms by showing that gene transfer between 'higher' eukaryotes can occur naturally. Analysis of the syntenic regions from several fishes strongly suggests that the smelt acquired the AFP gene from the herring.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".