Autotransporter-Encoding Sequences Are Phylogenetically Distributed among<i>Escherichia coli</i>Clinical Isolates and Reference Strains
Bibliographic record
Abstract
Autotransporters are secreted bacterial proteins exhibiting diverse virulence functions. Various autotransporters have been identified among Escherichia coli associated with intestinal or extraintestinal infections; however, the specific distribution of autotransporter sequences among a diversity of E. coli strains has not been investigated. We have validated the use of a multiplex PCR assay to screen for the presence of autotransporter sequences. Herein, we determined the presence of 13 autotransporter sequences and five allelic variants of antigen 43 (Ag43) among 491 E. coli isolates from human urinary tract infections, diarrheagenic E. coli, and avian pathogenic E. coli (APEC) and E. coli reference strains belonging to the ECOR collection. Clinical isolates were also classified into established phylogenetic groups. The results indicated that Ag43 alleles were significantly associated with clinical isolates (93%) compared to commensal isolates (56%) and that agn43K12 was the most common and widely distributed allele. agn43 allelic variants were also phylogenetically distributed. Sequences encoding espC, espP, and sepA and agn43 alleles EDL933 and RS218 were significantly associated with diarrheagenic E. coli strains compared to other groups. tsh was highly associated with APEC strains, whereas sat was absent from APEC. vat, sat, and pic were associated with urinary tract isolates and were identified predominantly in isolates belonging to either group B2 or D of the phylogenetic groups based on the ECOR strain collection. Overall, the results indicate that specific autotransporter sequences are associated with the source and/or phylogenetic background of strains and suggest that, in some cases, autotransporter gene profiles may be useful for comparative analysis of E. coli strains from clinical, food, and environmental sources.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".