Trimethoprim resistance in surface and wastewater is mediated by contrasting variants of the dfrB gene
Bibliographic record
Abstract
Trimethoprim (TMP) is a low-cost, widely prescribed antibiotic. Its effectiveness is increasingly challenged by the spread of genes coding for TMP-resistant dihydrofolate reductases: dfrA, and the lesser-known, evolutionarily unrelated dfrB. Despite recent reports of novel variants conferring high level TMP resistance (dfrB10 to dfrB21), the prevalence of dfrB is still unknown due to underreporting, heterogeneity of the analyzed genetic material in terms of isolation sources, and limited bioinformatic processing. In this study, we explored a coherent set of shotgun metagenomic sequences to quantitatively estimate the abundance of dfrB gene variants in aquatic environments. Specifically, we scanned sequences originating from influents and effluents of municipal sewage treatment plants as well as river-borne microbiomes. Our analyses reveal an increased prevalence of dfrB1, dfrB2, dfrB3, dfrB4, dfrB5, and dfrB7 in wastewater microbiomes as compared to freshwater. These gene variants were frequently found in genomic neighborship with other resistance genes, transposable elements, and integrons, indicating their mobility. By contrast, the relative abundances of the more recently discovered variants dfrB9, dfrB10, and dfrB13 were significantly higher in freshwater than in wastewater microbiomes. Moreover, their direct neighborship with other resistance genes or markers of mobile genetic elements was significantly less likely. Our findings suggest that natural freshwater communities form a major reservoir of the recently discovered dfrB gene variants. Their proliferation and mobilization in response to the exposure of freshwater communities to selective TMP concentrations may promote the prevalence of high-level TMP resistance and thus limit the future effectiveness of antimicrobial therapies.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.000 | 0.001 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".