ChemPLAN-Net: A deep learning framework to find novel inhibitor fragments for proteins
Suarez Vasquez, M. A.; Xue, M.; Lam, J. H.; Goonetilleke, E. C.; Gao, X.; Huang, X.
Show abstract
Fragment-based drug design plays an important role in the drug discovery process by reducing the complex small-molecule space into a more manageable fragment space. We leverage the power of deep learning to design ChemPLAN-Net; a model that incorporates the pairwise association of physicochemical features of both the protein drug targets and the inhibitor and learns from thousands of protein co-crystal structures in the PDB database to predict previously unseen inhibitor fragments. Our novel protocol handles the computationally challenging multi-label, multi-class problem, by defining a fragment database and using an iterative featurepair binary classification approach. By training ChemPLAN-Net on available co-crystal structures of the protease protein family, excluding HIV-1 protease as a target, we are able to outperform fragment docking and recover the targets inhibitor fragments found in co-crystal structures or identified by in-vitro cell assays.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Physics-inspired accuracy estimator for model-docked ligand complexes 97%
- BioStructNet: Structure-Based Network with Transfer Learning for Predicting Biocatalyst Functions 97%
- Learning a force field from small-molecule crystal lattice predictions enables consistent sub-Angstrom protein-ligand docking 96%
Similar papers in this journal
- AI-Augmented Physics-Based Docking for Antibody-Antigen Complex Prediction 96%
- Attention-based approach to predict drug-target interactions across seven target superfamilies 96%
- ProBASS: a language model with sequence and structural features for predicting the effect of mutations on binding affinity 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.