The evolutionary diversification of the Salmonella artAB toxin locus
Bibliographic record
Abstract
Salmonella enterica is a diverse species of bacterial pathogens comprised of >2,500 serovars with variable host ranges and virulence properties. Accumulating evidence indicates that two AB5-type toxins, typhoid toxin and ArtAB toxin, contribute to the more severe virulence properties of the Salmonella strains that encode them. It was recently discovered that there are two distinct types of artAB-like genetic elements in Salmonella: those that encode ArtAB toxins (artAB elements) and those in which the artA gene is degraded and the ArtB homolog, dubbed PltC, serves as an alternative delivery subunit for typhoid toxin (pltC elements). Here, we take a multifaceted approach to explore the evolutionary diversification of artAB-like genetic elements in Salmonella. We identify 7 subtypes of ArtAB toxins and 4 different PltC sequence groups that are distributed throughout the Salmonella genus. Both artAB and pltC are encoded within numerous diverse prophages, indicating a central role for phages in their evolutionary diversification. Genetic and structural analyses revealed features that distinguish pltC elements from artAB and identified evolutionary adaptations that enable PltC to efficiently engage typhoid toxin A subunits. For both pltC and artAB, we find that the sequences of the B subunits are especially variable, particularly amongst amino acid residues that fine tune the chemical environment of their glycan binding pockets. This study provides a framework to delineate the remarkably complex collection of Salmonella artAB/pltC-like genetic elements and provides a window into the mechanisms of evolution for AB5-type toxins.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".