Structural Studies of a Four-MBT Repeat Protein MBTD1
Bibliographic record
Abstract
BACKGROUND: The Polycomb group (PcG) of proteins is a family of important developmental regulators. The respective members function as large protein complexes involved in establishment and maintenance of transcriptional repression of developmental control genes. MBTD1, Malignant Brain Tumor domain-containing protein 1, is one such PcG protein. MBTD1 contains four MBT repeats. METHODOLOGY/PRINCIPAL FINDINGS: We have determined the crystal structure of MBTD1 (residues 130-566aa covering the 4 MBT repeats) at 2.5 A resolution by X-ray crystallography. The crystal structure of MBTD1 reveals its similarity to another four-MBT-repeat protein L3MBTL2, which binds lower methylated lysine histones. Fluorescence polarization experiments confirmed that MBTD1 preferentially binds mono- and di-methyllysine histone peptides, like L3MBTL1 and L3MBTL2. All known MBT-peptide complex structures characterized to date do not exhibit strong histone peptide sequence selectivity, and use a "cavity insertion recognition mode" to recognize the methylated lysine with the deeply buried methyl-lysine forming extensive interactions with the protein while the peptide residues flanking methyl-lysine forming very few contacts [1]. Nevertheless, our mutagenesis data based on L3MBTL1 suggested that the histone peptides could not bind to MBT repeats in any orientation. CONCLUSIONS: The four MBT repeats in MBTD1 exhibits an asymmetric rhomboid architecture. Like other MBT repeat proteins characterized so far, MBTD1 binds mono- or dimethylated lysine histones through one of its four MBT repeats utilizing a semi-aromatic cage. ENHANCED VERSION: This article can also be viewed as an enhanced version in which the text of the article is integrated with interactive 3D representations and animated transitions. Please note that a web plugin is required to access this enhanced functionality. Instructions for the installation and use of the web plugin are available in Text S1.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".