A New Family of Transcriptional Regulators Activating Biosynthetic Gene Clusters for Secondary Metabolites
Bibliographic record
Abstract
We previously identified the aur1 biosynthetic gene cluster (BGC) in Streptomyceslavendulae subsp. lavendulae CCM 3239 (formerly Streptomycesaureofaciens CCM 3239), which is responsible for the production of the unusual angucycline-like antibiotic auricin. Auricin is produced in a narrow interval of the growth phase after entering the stationary phase, after which it is degraded due to its instability at the high pH values reached after the production phase. The complex regulation of auricin BGC is responsible for this specific production by several regulators, including the key activator Aur1P, which belongs to the family of atypical response regulators. The aur1P gene forms an operon with the downstream aur1O gene, which encodes an unknown protein without any conserved domain. Homologous aur1O genes have been found in several BGCs, which are mainly responsible for the production of angucycline antibiotics. Deletion of the aur1O gene led to a dramatic reduction in auricin production. Transcription from the previously characterized Aur1P-dependent biosynthetic aur1Ap promoter was similarly reduced in the S. lavendulaeaur1O mutant strain. The aur1O-specific coactivation of the aur1Ap promoter was demonstrated in a heterologous system using a luciferase reporter gene. In addition, the interaction between Aur1O and Aur1P has been demonstrated by a bacterial two-hybrid system. These results suggest that Aur1O is a specific coactivator of this key auricin-specific positive regulator Aur1P. Bioinformatics analysis of Aur1O and its homologues in other BGCs revealed that they represent a new family of transcriptional coactivators involved in the regulation of secondary metabolite biosynthesis. However, they are divided into two distinct sequence-specific subclasses, each of which is likely to interact with a different family of positive regulators.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".