Genomic comparison of Neodiprion sertifer and Neodiprion lecontei nucleopolyhedroviruses and identification of potential hymenopteran baculovirus-specific open reading frames
Bibliographic record
Abstract
Genomic comparison of Neodiprion sertifer nucleopolyhedrovirus (NeseNPV) and Neodiprion lecontei nucleopolyhedrovirus (NeleNPV) showed that the hymenopteran baculoviruses had features in common and were distinct from other, fully sequenced lepidopteran and dipteran baculoviruses. Their genomes were small in size (86,462 and 81,755 bp, respectively), had low G+C contents (33.8 and 33.3 mol%, respectively) and contained fewer open reading frames (ORFs) (90 and 89, respectively) than other baculoviruses. They shared 69 ORFs (48.6% mean amino acid identity overall), 43 of which were previously identified baculovirus homologues. The remaining shared ORFs could be common to other baculoviruses, but low amino acid identities precluded identifying them as such. Some may also be unique to hymenopteran baculoviruses. These included a trypsin-like protease, a zinc-finger protein, regulator of chromosome condensation proteins, a densovirus capsid-like protein and a phosphotransferase. Structural analysis, the presence of conserved domains and phylogenetic studies suggested that some of these ORFs may be functional and could have been transferred horizontally from an insect host. ORFs found only in NeseNPV and NeleNPV may play a role in host specificity and/or tissue tropism, as hymenopteran baculoviruses are restricted to the midgut. The genomes were basically collinear, but contained non-syntenic regions (NSRs) with large numbers of repeats between their polyhedrin and dbp genes. They differed from each other in the number of ORFs and the G+C content of their NSRs and the presence of homologous regions in the NeseNPV genome. NeleNPV also had a short inversion relative to NeseNPV. NeseNPV contained 21 ORFs not found in NeleNPV and NeleNPV had 20 ORFs not found in NeseNPV.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".