Complete Sequence and Evolutionary Genomic Analysis of the <i>Pseudomonas aeruginosa</i> Transposable Bacteriophage D3112
Bibliographic record
Abstract
Bacteriophage D3112 represents one of two distinct groups of transposable phage found in the clinically relevant, opportunistic pathogen Pseudomonas aeruginosa. To further our understanding of transposable phage in P. aeruginosa, we have sequenced the complete genome of D3112. The genome is 37,611 bp, with an overall G+C content of 65%. We have identified 53 potential open reading frames, including three genes (the c repressor gene and early genes A and B) that have been previously characterized and sequenced. The organization of the putative coding regions corresponds to published genetic and transcriptional maps and is very similar to that of enterobacteriophage Mu. In contrast, the International Committee on Taxonomy of Viruses has classified D3112 as a lambda-like phage on the basis of its morphology. Similarity-based analyses identified 27 open reading frames with significant matches to proteins in the NCBI databases. Forty-eight percent of these were similar to Mu-like phage and prophage sequences, including proteins responsible for transposition, transcriptional regulation, virion morphogenesis, and capsid formation. The tail proteins were highly similar to prophage sequences in Escherichia coli and phage Phi12 from Staphylococcus aureus, while proteins at the right end were highly similar to proteins in Xylella fastidiosa. We performed phylogenetic analyses to understand the evolutionary relationships of D3112 with respect to Mu-like versus lambda-like bacteriophages. Different results were obtained from similarity-based versus phylogenetic analyses in some instances. Overall, our findings reveal a highly mosaic structure and suggest that extensive horizontal exchange of genetic material played an important role in the evolution of D3112.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".