Predicting Mutation-Induced Relative Protein-Ligand Binding Affinity Changes via Conformational Sampling and Diversity Integration with Subsampled Alphafold2 in Few-Shot Learning
Xie, L.; Lu, X.; Zhang, D.; Xu, L.; Yang, M.; Jeong, Y.-S.; Chang, S.; Xu, X.
Show abstract
Predicting the impact of mutations on protein-ligand binding affinity is crucial in drug discovery, particularly in addressing drug resistance and repurposing existing drugs. Current structure-based methods are constrained by their dependence on known protein-ligand complex structures. The absence of structural data for mutated proteins, along with the potential of mutations to modify protein conformational states, brings challenges for reliable predictions. The downstream application of conformational sampling of Alphafold, such as the prediction of binding free energy changes induced by mutation, has not been fully explored. To tackle the challenge, we propose a few-shot learning approach that integrates AlphaFold2 subsampling, ensemble docking, and Siamese learning to predict the relative binding affinity changes induced by mutations. We validate our approach for predicting binding affinity changes in 31 variants of the mutated Abelson tyrosine kinase (ABL). The conformation ensembles for each ABL variant are generated using AlphaFold2 subsampling. The most likely conformational states are selected by Gaussian fitting to construct the pairwise structural inputs and compute relative binding energies. Then the relative binding affinities are derived by averaging the values predicted by the Siamese learning model across the constructed conformational pairs. By benchmarking on the Tyrosine kinase inhibitors (TKI) dataset and the direct binding affinity using the refined set of PDBbind database, our proposed approach achieves an increased Spearman correlation coefficient for five out of six TKI molecules across 31 mutants when estimating relative binding affinities. Although the proposed method is validated on the ABL, it can be transferred and applied to other drug-target interaction predictions when the conformational flexibility of proteins needs to be considered. We anticipate the approach will be valuable for predicting binding energy changes induced by mutations in proteins or ligands, facilitating drug discovery efforts.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- CANDOCK: Chemical atomic network based hierarchical flexible docking algorithm using generalized statistical potentials 97%
- ProAffinity-GNN: A Novel Approach to Structure-based Protein-Protein Binding Affinity Prediction via a Curated Dataset and Graph Neural Networks 97%
- Activity Map and Transition Pathways of G Protein Coupled Receptor Revealed by Machine Learning 97%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Attention-based approach to predict drug-target interactions across seven target superfamilies 95%
- OPUS-X: An Open-Source Toolkit for Protein Torsion Angles, Secondary Structure, Solvent Accessibility, Contact Map Predictions, and 3D Folding 95%
- DrugTar Improves Druggability Prediction by Integrating Large Language Models and Gene Ontologies 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.