Engineered SH2 domains with tailored specificities and enhanced affinities for phosphoproteome analysis
Bibliographic record
Abstract
Protein phosphorylation is the most abundant post-translational modification in cells. Src homology 2 (SH2) domains specifically recognize phosphorylated tyrosine (pTyr) residues to mediate signaling cascades. A conserved pocket in the SH2 domain binds the pTyr side chain and the EF and BG loops determine binding specificity. By using large phage-displayed libraries, we engineered the EF and BG loops of the Fyn SH2 domain to alter specificity. Engineered SH2 variants exhibited distinct specificity profiles and were able to bind pTyr sites on the epidermal growth factor receptor, which were not recognized by the wild-type Fyn SH2 domain. Furthermore, mass spectrometry showed that SH2 variants with additional mutations in the pTyr-binding pocket that enhanced affinity were highly effective for enrichment of diverse pTyr peptides within the human proteome. These results showed that engineering of the EF and BG loops could be used to tailor SH2 domain specificity, and SH2 variants with diverse specificities and high affinities for pTyr residues enabled more comprehensive analysis of the human phosphoproteome. STATEMENT: Src Homology 2 (SH2) domains are modular domains that recognize phosphorylated tyrosine embedded in proteins, transducing these post-translational modifications into cellular responses. Here we used phage display to engineer hundreds of SH2 domain variants with altered binding specificities and enhanced affinities, which enabled efficient and differential enrichment of the human phosphoproteome for analysis by mass spectrometry. These engineered SH2 domain variants will be useful tools for elucidating the molecular determinants governing SH2 domains binding specificity and for enhancing analysis and understanding of the human phosphoproteome.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".