Back

Target-specific design of drug-like PPI inhibitors via hot-spot-guided generative deep learning

Sun, H.; Li, J.; Zhang, Y.; Lin, S.; Chen, J.; Tan, H.; Wang, R.; Mao, X.; Zhao, J.; Li, R.; Xiong, Y.; Wei, D.-Q.

2024-11-03 bioinformatics
10.1101/2024.10.29.620869 bioRxiv
Show abstract

Protein-protein interactions (PPIs) are vital therapeutic targets. However, the large and flat PPI interfaces pose challenges for the development of small-molecule inhibitors. Traditional computer-aided drug design approaches typically rely on pre-existing libraries or expert knowledge, limiting the exploration of novel chemical spaces needed for effective PPI inhibition. To overcome these limitations, we introduce Hot2Mol, a deep learning framework for the de novo design of drug-like, target-specific PPI inhibitors. Hot2Mol generates small molecules by mimicking the pharmacophoric features of hot-spot residues, enabling precise targeting of PPI interfaces without the need for bioactive ligands. The framework integrates three key components: a conditional transformer for pharmacophore-guided, drug-likeness-constrained molecular generation; an E(n)-equivariant graph neural network for accurate alignment with PPI hot-spot pharmacophores; and a variational autoencoder for generating novel and diverse molecular structures. Experimental evaluations demonstrate that Hot2Mol outperforms baseline models across multiple metrics, including docking affinities, drug-likenesses, synthetic accessibility, validity, uniqueness, and novelty. Furthermore, molecular dynamics simulations confirm the good binding stability of the generated molecules. Case studies underscore Hot2Mols ability to design high-affinity and selective PPI inhibitors, demonstrating its potential to accelerate rational PPI drug discovery.

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.