TMvisDB: resource for transmembrane protein annotation and 3D visualization
Marquet, C.; Grekova, A.; Houri, L.; Bernhofer, M.; Jimenez-Soto, L.; Karl, T.; Heinzinger, M.; Dallago, C.; Rost, B.
Show abstract
Since the rise of cellular organisms, transmembrane proteins (TMPs) have been crucial to a variety of cellular processes due to their central role as gates and gatekeepers. Despite their importance, experimental high-resolution structures for TMPs remain underrepresented due to technical limitations. With structure prediction methods coming of age, predictions might fill some of the need. However, identifying the membrane regions and topology in three-dimensional structure files requires additional in silico prediction. Here, we introduce TMvisDB to sieve through millions of predicted structures for TMPs. This resource enables both, to browse through 46 million predicted TMPs and to visualize those along with their topological annotations. The database was created by joining AlphaFold DB structure predictions and transmembrane topology predictions from the protein language model based method TMbed. We show the utility of TMvisDB for individual proteins through two single use cases, namely the B-lymphocyte antigen CD20 (Homo sapiens) and the cellulose synthase (Novosphingobium sp. P6W). To demonstrate the value for large scale analyses, we focus on all TMPs predicted for the human proteome. TMvisDB is freely available at tmvis.predictprotein.org.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- dagLogo: an R/Bioconductor package for identifying and visualizing differential amino acid group usage in proteomics data 95%
- AutoPhy: Automated phylogenetic identification of novel protein subfamilies 94%
- Redefining the architecture of ferlin proteins: insights into multi-domain protein structure and function 94%
Similar papers in this journal
- Comprehensive collection and prediction of ABC transmembrane protein structures in the AI era of structural biology 96%
- AutoCoEv -- a high-throughput in silico pipeline for predicting inter-protein co-evolution 95%
- In-Pero: Exploiting deep learning embeddings of protein sequences to predict the localisation of peroxisomal proteins 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.