Back

Sequence-independent protein domain detection and classification with PRISM

Tan, A.; Seedorf, H.

2026-05-27 bioinformatics
10.64898/2026.05.22.727336 bioRxiv
Show abstract

The explosion of predicted protein structures has revealed countless novel domain families. However, gold-standard segmentation tools like Chainsaw and Merizo are trained on rapidly obsoleting CATH databases, lack automatic domain classification, and cannot be easily fine-tuned without deep learning expertise. We introduce PRISM, a unified framework enabling sequence-independent, one-shot fine-tuning for simultaneous domain segmentation and classification, bypassing traditional constraints to accurately resolve complex, novel protein architectures.

Matching journals

The top 1 journal accounts for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.