Deep-learning predictions of biomolecular structures : persistent limitations and new horizons extended by explicit ion addition
Marien, J.; Sritharan, S.; Caviglia, B.; Versini, R.; Auclair, L.; Barraud, P.; Tisne, C.; Reguei, A.; Murail, S.; Leclerc, F.; Duboue-Dijon, E.; Tao, J.; Basdevant, N.; Baaden, M.; Prevost, C.; Sacquin-Mora, S.; Taly, A.
Show abstract
The advent of deep learning-driven tools such as AlphaFold has revolutionized the prediction of biomolecular structures, offering unprecedented accuracy and accessibility for proteins, RNA, and their complexes. While these tools have demonstrated remarkable success in benchmarking competitions and enabled experimentalists to generate models with ease, their widespread use has also highlighted persistent challenges. These include difficulties in assessing model confidence, limitations in predicting transmembrane domains, nucleic acids, conformational diversity, and interactions with ions or ligands, as well as the tendency to misfold intrinsically disordered regions (IDRs). In this perspective, we critically evaluate the strengths and limitations of current AI-based structure prediction tools through illustrative examples with a particular emphasis on the impact of explicit ion modelling. We notably report how the explicit addition of a few potassium cations to the prediction of IDRs or G-quadruplexes can trigger massive conformational switches compared to "dry" predictions. On this basis, we suggest modelling sequences both "dry" and in the presence of explicit potassium cations as a simple, practical way to sample alternative conformations and to expose disordered regions that current predictors tend to over-fold. We discuss the importance of reporting confidence metrics in publications to avoid overinterpretation. Furthermore, we address the unique challenges of RNA structure prediction, where data scarcity and structural complexity limit the performance of both classical and deep learning methods. Our analysis underscores the need for continued methodological advancements, integration of complementary computational tools, and expansion of high-quality experimental datasets. TOC Graphic O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=111 SRC="FIGDIR/small/740587v1_ufig1.gif" ALT="Figure 1"> View larger version (16K): org.highwire.dtl.DTLVardef@fc2f42org.highwire.dtl.DTLVardef@82c1a4org.highwire.dtl.DTLVardef@7733cborg.highwire.dtl.DTLVardef@1e97db2_HPS_FORMAT_FIGEXP M_FIG C_FIG
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- RosENet: Improving binding affinity prediction by leveraging molecular mechanics energies with a 3D Convolutional Neural Network 97%
- Influence of stereochemistry in a local approach for calculating protein conformations 97%
- Are Deep Learning Structural Models Sufficiently Accurate for Free Energy Calculations? Application of FEP+ to AlphaFold2 Predicted Structures 97%
Similar papers in this journal
- Multi-state Modeling of G-protein Coupled Receptors at Experimental Accuracy 96%
- Do Newly Born Orphan Proteins Resemble Never Born Proteins? A Study Using Three Deep Learning Algorithms 96%
- Novel sampling strategies and a coarse-grained score function for docking homomers, flexible heteromers, and oligosaccharides using Rosetta in CAPRI Rounds 37-45 96%
Similar papers in this journal
- RosettaDDGPrediction for high-throughput mutational scans: from stability to binding 97%
- ExploreTurns: A web tool for the exploration, analysis, and classification of beta turns and structured loops in proteins; application to beta-bulge and Schellman loops, Asx helix caps, beta hairpins and other hydrogen-bonded motifs 97%
- Knot or Not? Sequence-Based Identification of Knotted Proteins With Machine Learning 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.