MétaCan
Menu
Back to cohort
Record W2128588041 · doi:10.1002/prot.10424

NMR structure of the hypothetical protein AQ‐1857 encoded by the Y157 gene from <i>Aquifex aeolicus</i> reveals a novel protein fold

2004· article· en· W2128588041 on OpenAlexaff
Duanxiang Xu, Gaohua Liu, Rong Xiao, T.B. Acton, Sharon Goldsmith‐Fischman, Barry Honig, G.T. Montelione, Thomas Szyperski

Bibliographic record

VenueProteins Structure Function and Bioinformatics · 2004
Typearticle
Languageen
FieldBiochemistry, Genetics and Molecular Biology
TopicBacterial Genetics and Biotechnology
Canadian institutionsStructural Genomics Consortium
FundersNational Institute of General Medical SciencesNational Institutes of Health
KeywordsAquifex aeolicusChemistrySize-exclusion chromatographyEscherichia coliChromatographyGeneBiochemistry

Abstract

fetched live from OpenAlex

Uniformly (U) 13C, 15N-labeled AQ-1857 was cloned, expressed and purified following standard protocols. Briefly, the full length gene (Y157_AQUAE) from Aquifex aeolicus was cloned into a pET21d (Novagen) derivative, yielding the plasmid pQR6-21. The resulting construct contains eight nonnative residues at the C-terminus (LEHHHHHH) that facilitate protein purification. Escherichia coli BL21 (DE3) pMGK cells, a rare codon enhanced strain, were transformed with pQR6-21, and cultured in MJ minimal medium containing (15NH4)2SO4 and U-13C-glucose as sole nitrogen and carbon sources. U-13C,15N AQ-1857 was purified using a two-step protocol consisting of Ni-NTA affinity (Qiagen) and gel filtration (HiLoad 26/60 Superdex 75, Amersham Biosciences) chromatography. The final yield of purified U-13C, 15N AQ-1857 (> 97% homogeneous by SDS-PAGE; 14.4 kDa by MALDI-TOF mass spectrometry) was about 10 mg/L. In addition, a sample which was U-15N and 5% biosynthetically directed fractionally 13C-labeled was generated for stereospecific assignment of isopropyl methyl groups.8 Two samples of 5%13C,U-15N and U-13C,15N AQ-1857 were prepared at concentrations of 1.0 mM in 95% H2O/5% D2O solution containing 20 mM MES, 100 mM NaCl, 10 mM DTT, 5 mM CaCl2, 0.02% NaN3 at pH 6.5. All NMR data were collected at 20°C on Varian INOVA 600 and 750 spectrometers. The spectra were processed and analyzed using the programs NMRPipe9 and XEASY,10 respectively. Resonance assignments were obtained as described11 using a suite of reduced-dimensionality NMR experiments, including 3D HNNCAHA, HαβCαβ(CO)NHN, HCCH-COSY, and 2D HBCB(CGCD)HD. These data were complemented by conventional12 HNNCACB and HC(C)H TOCSY experiments. Assignments were obtained for 93% of the backbone and 13Cβ, and for 91% of the side chain chemical shifts. Stereospecific assignments were obtained for 44% of the β-methylene groups exhibiting non-degenerate proton chemical shifts, and for all Val and Leu isopropyl moieties. The chemical shifts were deposited in the BioMagResBank (accession code: 5683). Upper distance limit constraints for structure calculations were obtained from 3D 15N-and 13C-resolved [1H,1H]-NOESY12 (Table I). In addition, 3JHNα scalar couplings measured in 3D HNNHA12 yielded ϕ-angle constraints, and backbone dihedral angle constraints were derived from chemical shifts as described13 for residues located in regular secondary structure elements (Table I). Structure calculations were performed using the program DYANA.14 Statistics for the structure determination (Table I) show that a high-quality NMR structure was obtained (Fig. 1). AQ-1857 (PDB ID: 1NWB) contains seven β-strands A to F and two α-helices. A(↓), F(↓) and G(↑) form a 3-stranded, and D(↓), E(↑), B(↑) and C(↓) form a 4-stranded sheet. The two sheets form a “sandwich” being rotated by ∼45 degrees relative to each other (Fig. 2). The segment 40–45 and the C-terminal tail 102–116 are flexibly disordered in solution. The 20 DYANA conformers with the lowest residual DYANA target function chosen to represent the NMR solution structure of AQ-1857 are shown after superposition of the backbone heavy atoms N, Cα and C′ of the regular secondary structure elements for minimal RMSD. A: The novel fold of AQ-1857: ribbon drawing of the DYANA conformer with the lowest residual target function value. The α-helices I and II are shown in red and yellow, the β-strands A to G are in cyan, other polypeptide segments are in grey, and the N- and C-terminal ends of the protein are indicated as ‘N’ and ‘C’. B: Same as in (A), but rotated by 90° about the vertical axis. α-Helix I: residues 14–25, II: 77–79; β-Strand A: 10–12, B: 33–36, C: 52–53, D: 63–65, E: 69–72, F: 84–89, G: 94–99. The NMR structure of AQ-1857 is the first structure representative of the larger HesB family2, 3 of proteins. No meaningful structural homologues were identified using the programs SKAN,15 DALI,16 or CE.17 This finding strongly supports the notion that AQ-1857 possesses a hitherto uncharacterized, novel fold (Fig. 2). Interestingly, the PROSITE consensus pattern for the HesB family spans the flexibly disordered C-terminal tail of the protein. In fact, two of the three cysteinyl residues which have been proposed to be involved in iron-sulfur cluster assembly in members of this family5, 6 are located in this tail (not shown in Fig. 1; the third cysteine is located in position 43 in the loop connecting β-strands 2 and 3). It is thus very likely that the flexibly disordered tail is of functional importance, and it is tempting to speculate that this tail adopts an ordered conformation only upon involvement in Fe-S cluster assembly. This work was supported by the National Institutes of Health (P50 GM62413-01), the National Science Foundation (MCB 00075773 to T.S.; DBI-9904841 to B.H), and the Center for Computational Research at UB.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.000
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Bench or experimental · Consensus signal: Bench or experimental
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.014
Threshold uncertainty score0.828

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0000.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0000.000
Research integrity0.0010.000
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.005
GPT teacher head0.181
Teacher spread0.176 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designBench or experimental
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations8
Published2004
Admission routes1
Has abstractyes

Explore more

Same venueProteins Structure Function and BioinformaticsSame topicBacterial Genetics and BiotechnologyFrench-language works237,207