A Novel AT-Rich DNA Recognition Mechanism for Bacterial Xenogeneic Silencer MvaT
Bibliographic record
Abstract
Bacterial xenogeneic silencing proteins selectively bind to and silence expression from many AT rich regions of the chromosome. They serve as master regulators of horizontally acquired DNA, including a large number of virulence genes. To date, three distinct families of xenogeneic silencers have been identified: H-NS of Proteobacteria, Lsr2 of the Actinomycetes, and MvaT of Pseudomonas sp. Although H-NS and Lsr2 family proteins are structurally different, they all recognize the AT-rich DNA minor groove through a common AT-hook-like motif, which is absent in the MvaT family. Thus, the DNA binding mechanism of MvaT has not been determined. Here, we report the characteristics of DNA sequences targeted by MvaT with protein binding microarrays, which indicates that MvaT prefers binding flexible DNA sequences with multiple TpA steps. We demonstrate that there are clear differences in sequence preferences between MvaT and the other two xenogeneic silencer families. We also determined the structure of the DNA-binding domain of MvaT in complex with a high affinity DNA dodecamer using solution NMR. This is the first experimental structure of a xenogeneic silencer in complex with DNA, which reveals that MvaT recognizes the AT-rich DNA both through base readout by an "AT-pincer" motif inserted into the minor groove and through shape readout by multiple lysine side chains interacting with the DNA sugar-phosphate backbone. Mutations of key MvaT residues for DNA binding confirm their importance with both in vitro and in vivo assays. This novel DNA binding mode enables MvaT to better tolerate GC-base pair interruptions in the binding site and less prefer A tract DNA when compared to H-NS and Lsr2. Comparison of MvaT with other bacterial xenogeneic silencers provides a clear picture that nature has evolved unique solutions for different bacterial genera to distinguish foreign from self DNA.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".