Multilocus Sequence Typing Scheme for the Characterization of 936-Like Phages Infecting Lactococcus lactis
Bibliographic record
Abstract
Lactococcus lactis phage infections are costly for the dairy industry because they can slow down the fermentation process and adversely impact product safety and quality. Although many strategies have been developed to better control phage populations, new virulent phages continue to emerge. Thus, it is beneficial to develop an efficient method for the routine identification of new phages within a dairy plant to rapidly adapt antiphage tactics. Here, we present a multilocus sequence typing (MLST) scheme for the characterization of the 936-like phages, the most prevalent phage group infecting L. lactis strains worldwide. The proposed MLST system targets the internal portion of five highly conserved genomic sequences belonging to the packaging, morphogenesis, and lysis modules. Our MLST scheme was used to analyze 100 phages with different restriction fragment length polymorphism (RFLP) patterns isolated from 11 different countries between 1971 and 2010. PCR products were obtained for all the phages analyzed, and sequence analysis highlighted the high discriminatory power of the MLST system, detecting 93 different sequence types. A conserved locus within the lys gene (coding for endolysin) was the most discriminative, with 65 distinct alleles. The locus within the mcp gene (major capsid protein) was the most conserved (54 distinct alleles). Phylogenetic analyses of the concatenated sequences exhibited a strong concordance of the clusters with the phage host range, indicating the clonal evolution of these phages. A public database has been set up for the proposed MLST system, and it can be accessed at http://pubmlst.org/bacteriophages/.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".