Transcriptome analysis of neoplastic hemocytes in soft-shell clams Mya arenaria: Focus on cell cycle molecular mechanism
Bibliographic record
Abstract
In North America, a high mortality of soft-shell clams Mya arenaria was found to be related to the disease known as disseminated neoplasia (DN). Disseminated neoplasia is commonly recognized as a tetraploid disorder related to a disruption of the cell cycle. However, the molecular mechanisms by which hemocytes of clams are transformed in the course of DN remain by far unknown. This study aims at identifying the transcripts related to DN in soft shell clams' hemocytes using next generation of sequencing (Illumina HiSeq2000). This study mainly focuses on transcripts and molecular mechanisms involved in cell cycle. Using Illumina next generation of sequencing, more than 95,399,159 reads count with an average length of 45 bp was generated from three groups of hemocytes: (1) a healthy group with less than 10% of tetraploid cells; (2) an intermediate group with tetraploid hemocytes ranging between 10% and 50% and (3) a diseased group with more than 50% of tetraploid cells. After the reads were cleaned by removing the adapters, de novo assembly was performed on the sequences and more than 73,696 contigs were generated with a mean contig length estimated at 585 bp ranging from 189 bp to 14,773 bp. Once a Blastx search against NCBI Non Redundant database was performed and the duplicates removed, 18,378 annotated sequences matched known sequences, 3078 were hypothetical and 9002 were uncharacterized sequences. Fifty percent and 41% of known sequences match sequences from Mollusca and Gastropoda respectively. Among the bivalvia, 33%, 17%, 17% and 15% of the contigs match sequences from Ostreoida, Veneroida, Pectinoida and Mytiloida respectively. Gene ontology analysis showed that metabolic, cellular, transport, cell communication and cell cycle represent 33%, 15%, 9%, 8.5% and 7% respectively of the total biological process. Approximately 70% of the component process is related to intracellular process and 15% is linked to protein and ribonucleoprotein complex. Catalytic activities and binding molecular processes represent 39% and 33% of the total molecular functions. Interestingly, nucleic acid binding represents more than 18% of the total protein class. Transcripts involved in the molecular mechanisms of cell cycle are discussed providing new avenues for future investigations.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.005 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".