AlphaFold2 and AlphaFold3 leads to significantly different results in human-parasite interaction prediction
Ozden, B.; Cuesta Astroz, Y.; Karaca, E.
Show abstract
MotivationParasitic diseases pose a significant global health challenge with substantial socioeconomic impact, especially in developing countries (Dampier et al., 2009; Mehmood et al., 2017; Merrifield et al., 2016). Combatting these diseases is difficult due to two factors: the resistance developed by parasites to existing drugs and the underfunded research efforts to resolve host-parasite interactions (Cuesta-Astroz & Oliveira, 2018). To address the latter, we focused on 276 human-parasite domain-domain interaction predictions, involving 15 parasitic species known to cause infections in humans (Table S1, Cuesta-Astroz et al., 2019). We aimed to validate these interactions by using all AlphaFold2-Multimer (AF2) and AlphaFold3 (AF3) (Jumper et al., 2021; Abramson et al., 2024). Our initial hypothesis was that the interactions with an AF confidence score [≥] 0.8 could serve as high confident templates for further research. However, through systematic modeling of these interactions, we observed a striking discrepancy: AF3 confidence assignment behaves significantly differently from that of AF2. In this study, we investigate the underlying causes of this difference.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Towards a comprehensive view of the pocketome universe - biological implications and algorithmic challenges. 95%
- Improved protein complex prediction with AlphaFold-multimer by denoising the MSA profile 94%
- ECOD domain classification of 48 whole proteomes from AlphaFold Structure Database using DPAM 93%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.