FlashFold: a standalone command-line tool for accelerated protein structure prediction
Saha, C. K.; Roghanian, M.; Häussler, S.; Guy, L.
Show abstract
AlphaFold has revolutionized the decades-old issue of precisely predicting protein structures. However, its high accuracy relies on a computationally intensive step that involves searching vast databases for homologous sequences as the query protein of interest. Additionally, predicting the quaternary structure of protein complexes requires prior knowledge of subunit counts, a prerequisite rarely met. To address these limitations, we introduce FlashFold - a fast, user-friendly tool for protein structure prediction. It accelerates homology searches using a compact built-in database, enabling structure predictions up to 3-fold faster than AlphaFold3, with sacrificing little or no accuracy. Unlike others, FlashFold features adaptable built-in databases that allow users to easily incorporate their own privately sequenced data - an option that can positively influence the prediction accuracy. Moreover, it allows users to estimate stoichiometry of protein complexes directly from sequence information. To support high-throughput workflows and streamline downstream decision-making, it generates interactive and filterable summary reports, enabling users to efficiently visualize protein structures, and interpret large volumes of prediction results. FlashFold runs locally on Linux and Mac, eliminating reliance on third-party servers. FlashFold is available at https://github.com/chayan7/flashfold.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- SOLeNNoID: A Deep Learning Pipeline For Solenoid Residue Detection in Protein Structures 97%
- CATHe: Detection of remote homologues for CATH superfamilies using embeddings from protein language models 96%
- CONSTRUCT: an algorithmic tool for identifying functional or structurally important regions in protein tertiary structure 96%
Similar papers in this journal
- The Protein Common Assembly Database (ProtCAD): A comprehensive structural resource of protein complexes 96%
- Kincore: a web resource for structural classification of protein kinases and their inhibitors 95%
- doubleHelix: nucleic acid sequence identification, assignment and validation tool for cryo-EM and crystal structure models 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.