Back

SENSE-PPI reconstructs protein-protein interactions of various complexities, within, across, and between species, with sequence-based evolutionary scale modeling and deep learning

Volzhenin, K.; Bittner, L.; Carbone, A.

2023-10-12 bioinformatics
10.1101/2023.09.19.558413 bioRxiv
Show abstract

Ab initio computational reconstructions of protein-protein interaction (PPI) networks will provide invaluable insights on cellular systems, enabling the discovery of novel molecular interactions and elucidating biological mechanisms within and between organisms. Leveraging latest generation protein language models and recurrent neural networks, we present SENSE-PPI, a sequence-based deep learning model that efficiently reconstructs ab initio PPIs, distinguishing partners among tens of thousands of proteins and identifying specific interactions within functionally similar proteins. SENSE-PPI demonstrates high accuracy, limited training requirements, and versatility in cross-species predictions, even with non-model organisms and human-virus interactions. Its performance decreases for phylogenetically more distant model and non-model organisms, but signal alteration is very slow. SENSE-PPI is state-of-the-art, outperforming all existing methods. In this regard, it demonstrates the important role of parameters in protein language models. SENSE-PPI is very fast and can test 10,000 proteins against themselves in a matter of hours, enabling the reconstruction of genome-wide proteomes. Graphical abstractSENSE-PPI is a general deep learning architecture predicting protein-protein interactions of different complexities, between stable proteins, between stable and intrinsically disordered proteins, within a species, and between species. Trained on one species, it accurately predicts interactions and reconstructs complete specialized subnetworks for model and non-model organisms, and trained on human-virus interactions, it predicts human-virus interactions for new viruses. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=187 SRC="FIGDIR/small/558413v3_ufig1.gif" ALT="Figure 1"> View larger version (63K): org.highwire.dtl.DTLVardef@106d345org.highwire.dtl.DTLVardef@11883a6org.highwire.dtl.DTLVardef@6b07d9org.highwire.dtl.DTLVardef@d07138_HPS_FORMAT_FIGEXP M_FIG C_FIG

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.