Diagnosing protein sequence search in the era of language models
Zhou, H.; Yang, Y.; Lu, Y. Y.
Show abstract
Protein language model (PLM) based search is rapidly emerging as a successor to classical sequence alignment, with recent high-profile studies reporting substantial improvements in speed and remote homology detection. However, success on standard benchmarks does not guarantee that similarity derived from PLM embeddings constitutes reliable biological evidence. Here, we introduce PLM-GUARD, a diagnostic framework designed to interrogate the underlying meaning of protein search scores and assess their biological trustworthiness. PLM-GUARD comprises six sanity checks spanning biological fidelity, semantic validity, and manipulation safety. Across eight representative search methods, classical alignment-based systems demonstrate remarkable robustness, whereas current PLM-based methods fail broadly across all three dimensions. Notably, hybrid methods show intermediate results, indicating that alignment is still critical for ensuring biologically grounded correspondence. Our findings provide a timely clarification for the field and underscore the necessity of diagnostic evaluation as protein search enters the era of language models.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Sequence-based prediction of protein-protein interactions: a structure-aware interpretable deep learning model 96%
- DynamicGT: a dynamic-aware geometric transformer model to predict protein binding interfaces in flexible and disordered regions 95%
- Undersampling and the inference of coevolution in proteins 95%
Similar papers in this journal
- Controllable Protein Design via Autoregressive Direct Coupling Analysis Conditioned on Principal Components 96%
- Phylogenetic correlations can suffice to infer protein partners from sequences 95%
- Paraplume: A fast and accurate paratope prediction method provides insights into repertoire-scale binding dynamics 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.