The transcriptomics of sympatric dwarf and normal lake whitefish (Coregonus clupeaformis spp., Salmonidae) divergence as revealed by next-generation sequencing
Bibliographic record
Abstract
Gene expression divergence is one of the mechanisms thought to be involved in the emergence of incipient species. Next-generation sequencing has become an extremely valuable tool for the study of this process by allowing whole transcriptome sequencing, or RNA-Seq. We have conducted a 454 GS-FLX pyrosequencing experiment to refine our understanding of adaptive divergence between dwarf and normal lake whitefish species (Coregonus clupeaformis spp.). The objectives were to: (i) investigate transcriptomic divergence as measured by liver RNA-Seq; (ii) test the correlation between divergence in expression and sequence polymorphism; and (iii) investigate the extent of allelic imbalance. We also compared the results of RNA-seq with those of a previous microarray study performed on the same fish. Following de novo assembly, results showed that normal whitefish overexpressed more contigs associated with protein synthesis while dwarf fish overexpressed more contigs related to energy metabolism, immunity and DNA replication and repair. Moreover, 63 SNPs showed significant allelic imbalance, and this phenomenon prevailed in the recently diverged dwarf whitefish. Results also showed an absence of correlation between gene expression divergence as measured by RNA-Seq and either polymorphism rate or sequence divergence between normal and dwarf whitefish. This study reiterates an important role for gene expression divergence, and provides evidence for allele-specific expression divergence as well as evolutionary decoupling of regulatory and coding sequences in the adaptive divergence of normal and dwarf whitefish. It also demonstrates how next-generation sequencing can lead to a more comprehensive understanding of transcriptomic divergence in a young species pair.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".