Crystal structure of<i>Saccharomyces cerevisiae</i>homologous mitochondrial matrix factor 1 (Hmf1)
Bibliographic record
Abstract
We have initiated a structural genomics project based on selected protein families, with the representatives for structural studies obtained from the genome of E. coli. A total of 160 genes have been cloned to date, with 126 successfully overexpressing soluble protein with at least one fusion system. We are currently applying robotics methods allowing cloning in an automated manner. Target genes have been cloned as N-terminal fusions with GST, (His)6, or (His)8 affinity tags. Over 60 proteins have so far been purified to homogeneity. Purified proteins are further characterized for homogeneity and suitable solution properties using a combination of dynamic light scattering and electrophoretic methods. Purified protein samples are screened for initial crystallization conditions in 96-well format using a sparse-matrix approach. A specialized database has been developed to follow the progress of various proteins and to store all relevant experimental data (http://sgen.bri.nrc.ca/bsgi). Web-based software for displaying and searching the information from this local database and for tracking deposited structures in the PDB as well as the progress of other structural genomics projects has been developed. To date, crystals have been obtained for 37 of the purified proteins. of these, diffraction quality crystals were obtained for 16 proteins. Using SeMet-substituted proteins we have determined the structures of 13 of these proteins by MAD phasing. Structures determined include 2amino-3-ketobutyrate CoA ligase, part of the threonine salvage pathway, MoeA protein, involved in molydopterin biosynthesis, histidinol phosphate aminotransferase and L-histidinol dehydrogenase, two enzymes associated with histidine biosynthesis, and RsuA, a 16S rRNA psuedouridine synthase.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.001 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".