Assessing hERG1 Blockade from Bayesian Machine-Learning-Optimized Site Identification by Ligand Competitive Saturation Simulations
Bibliographic record
Abstract
-Related (hERG1) potassium channel. There is a compulsory preclinical stage safety assessment for the hERG1 blockade for all classes of drugs, which adds substantially to the cost of drug development. The availability of a high-resolution cryogenic electron microscopy (cryo-EM) structure for the channel in its open/depolarized state solved in 2017 enabled the application of molecular modeling for rapid assessment of drug blockade by molecular docking and simulation techniques. More importantly, if successful, in silico methods may allow a path to lead-compound salvaging by mapping out key block determinants. Here, we report the blind application of the site identification by the ligand competitive saturation (SILCS) protocol to map out druggable/regulatory hotspots in the hERG1 channel available for blockers and activators. The SILCS simulations use small solutes representative of common functional groups to sample the chemical space for the entire protein and its environment using all-atom simulations. The resulting chemical maps, FragMaps, explicitly account for receptor flexibility, protein-fragment interactions, and fragment desolvation penalty allowing for rapid ranking of potential ligands as blockers or nonblockers of hERG1. To illustrate the power of the approach, SILCS was applied to a test set of 55 blockers with diverse chemical scaffolds and pIC50 values measured under uniform conditions. The original SILCS model was based on the all-atom modeling of the hERG1 channel in an explicit lipid bilayer and was further augmented with a Bayesian-optimization/machine-learning (BML) stage employing an independent literature-derived training set of 163 molecules. BML approach was used to determine weighting factors for the FragMaps contributions to the scoring function. pIC50 predictions from the combined SILCS/BML approach to the 55 blockers showed a Pearson correlation (PC) coefficient of >0.535 relative to the experimental data. SILCS/BML model was shown to yield substantially improved performance as compared to commonly used rigid and flexible molecular docking methods for a well-established cohort of hERG1 blockers, where no correlation with experimental data was recorded. SILCS/BML results also suggest that a proper weighting of protonation states of common blockers present at physiological pH is essential for accurate predictions of blocker potency. The precalculated and optimized SILCS FragMaps can now be used for the rapid screening of small molecules for their cardiotoxic potential as well as for exploring alternative binding pockets in the hERG1 channel with applications to the rational design of activators.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.005 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".