Protein structure prediction in the era of AI: challenges and limitations when applying to in-silico force spectroscopy
Gomes, P. S. F. C.; Gomes, D. E. B.; Bernardi, R. C.
Show abstract
Mechanoactive proteins are essential for a myriad of physiological and pathological processes. Guided by the advances in single-molecule force spectroscopy (SMFS), we have reached a molecular-level understanding of how several mechanoactive proteins respond to mechanical forces. However, even SMFS has its limitations, including the lack of detailed structural information during force-loading experiments. That is where molecular dynamics (MD) methods shine, bringing atomistic details with femtosecond time-resolution. However, MD heavily relies on the availability of high-resolution structures, which is not available for most proteins. For instance, the Protein Data Bank currently has 192K structures deposited, against 231M protein sequences available on Uniprot. But many are betting that this gap might become much smaller soon. Over the past year, the AI-based AlphaFold created a buzz on the structural biology field by being able to, for the first time, predict near-native protein folds from their sequences. For some, AlphaFold is causing the merge of structural biology with bioinformatics. In this perspective, using an in silico SMFS approach, we investigate how reliable AlphaFold structure predictions are to investigate mechanical properties of staph bacteria adhesins proteins. Our results show that AlphaFold produce extremally reliable protein folds, but in many cases is unable to predict high-resolution protein complexes accurately. Nonetheless, the results show that AlphaFold can revolutionize the investigation of these proteins, particularly by allowing high-throughput scanning of protein structures. Meanwhile, we show that the AlphaFold results need to be validated and should not be employed blindly, with the risk of obtaining an erroneous protein mechanism.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Do Newly Born Orphan Proteins Resemble Never Born Proteins? A Study Using Three Deep Learning Algorithms 96%
- Novel sampling strategies and a coarse-grained score function for docking homomers, flexible heteromers, and oligosaccharides using Rosetta in CAPRI Rounds 37-45 96%
- Molecular Dynamics Analysis of a Flexible Loop at the Binding Interface of the SARS-CoV-2 Spike Protein Receptor-Binding Domain 96%
Similar papers in this journal
- Exploring the ability of the MD+FoldX method to predict SARS-CoV-2 antibody escape mutations using large-scale data 96%
- Benchmark of force fields to characterize the intrinsically disordered R2-FUS-LC region 95%
- The Structural Basis of the Genetic Code: Amino Acid Recognition by Aminoacyl-tRNA Synthetases 95%
Similar papers in this journal
- The Impact of Gag Non-Cleavage Site Mutations on HIV-1 Viral Fitness from Integrative Modelling and Simulations 95%
- Leri: a web-server for identifying protein functional networks from evolutionary couplings 95%
- Systematic Investigation of Machine Learning on Limited Data: A Study on Predicting Protein-Protein Binding Strength 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.