gapTrick - Structural characterisation of protein-protein interactions using AlphaFold with multimeric templates
Chojnowski, G.
Show abstract
The structural characterisation of protein-protein interactions is a key step in understanding the functions of living cells. Here I show that AlphaFold3 often fails to predict protein complexes that are either weak, or dependent on the presence of a cofactor that is not included in a prediction. To address this problem, I developed gapTrick, an AlphaFold2-based approach that uses multimeric templates to improve prediction reliability. I show that gapTrick improves predictions of weak and incomplete complexes based on low-accuracy templates, such as separate protein models rigid-body fitted into a cryo-EM reconstruction. I also show that it identifies with very high precision residue-residue interactions, which are critical for complex formation and a very strong indicator of model correctness. The approach can aid in the interpretation of challenging experimental structures and the computational identification of protein-protein interactions. Availability and implementationThe gapTrick source code is available at https://github.com/gchojnowski/gapTrick and requires only a standard AlphaFold2 installation to run. The repository also provides a Colab notebook that can be used to run gapTrick without installing it on the users computer.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- CICLOP: A Robust, Faster, and Accurate Computational Framework for Protein Inner Cavity Detection 96%
- ProBASS: a language model with sequence and structural features for predicting the effect of mutations on binding affinity 95%
- SPAED: Harnessing AlphaFold Output for Accurate Segmentation of Phage Endolysin Domains 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.