Stabilizing interactions in the dimer interface of α‐subunit in <i>Escherichia coli</i> RNA polymerase: A graph spectral and point mutation study
Bibliographic record
Abstract
The formation of alpha(2) dimer in Escherichia coli core RNA polymerase (RNAP) is thought to be the first step toward the assembly of the functional enzyme. A large number of evidences indicate that the alpha-subunit dimerizes through its N-terminal domain (NTD). The crystal structures of the alpha-subunit NTD and that of a homologous Thermus aquaticus core RNAP are known. To identify the stabilizing interactions in the dimer interface of the alpha-NTD of E. coli RNAP, we identified side-chain clusters by using the crystal structure coordinates of E. coli alpha-NTD. A graph spectral algorithm was used to identify side-chain clusters. This algorithm considers the global nonbonded side-chain interactions of the residues for the clustering procedure and is unique in identifying residues that make the largest number of interactions among the residues that form clusters in a very quantitative way. By using this algorithm, a nine-residue cluster consisting of polar and hydrophobic residues was identified in the subunit interface adjacent to the hydrophobic core. The residues forming the cluster are relatively rigid regions of the interface, as measured by the thermal factors of the residues. Most of the cluster residues in the E. coli enzyme were topologically and sequentially conserved in the T. aquaticus RNAP crystal structure. Residues 35F and 46I were predicted to be important in the stability of the alpha-dimer interface, with 35F forming the center of the cluster. The predictions were tested by isolating single-point mutants alpha-F35A and alpha-I46S on the dimer interface, which were found to disrupt dimerization. Thus, the identified cluster at the edge of the dimer interface seems to be a vital component in stabilizing the alpha-NTD.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".