Back

Cross-Attention Over RNA And Protein Sequences Enables Generalizable Interaction Prediction

Catalano, M.; Pepe, G.; Ausiello, G.; McWhite, C.; Gambosi, G.; Helmer-Citterich, M.; Gherardini, P. F.

2026-04-23 bioinformatics
10.64898/2026.04.22.720174 bioRxiv
Show abstract

Computational predictions are essential to characterize the RNA-protein interaction landscape, yet a persistent gap between benchmark performance and practical utility suggests that current models have limited generalization capabilities. To address this issue, we present CORAL (Cross-attention for RNA-protein Association Learning), a deep learning framework for the prediction of RNA-protein interactions that integrates pretrained protein (ESM-2) and RNA (DNABERT2) language models through bidirectional cross-attention with Low-Rank Adaptation fine-tuning. We also introduce a benchmarking framework that rigorously addresses the problem of data redundancy between training and test sets, which greatly inflates model performances reported in the literature. To this end we adopt three partitioning strategies of increasing stringency: conventional random splits, pairwise non-redundant splits, and component-wise non-redundant splits. CORAL maintains an F1 score of 0.65 under the most stringent component-wise evaluation, compared to 0.55 for the next-best method. Interpretability analyses reveal that specific cross-attention heads systematically attend to structurally defined contact positions between RNA and protein molecules, showing 27% elevated attention at interface residues across 309 experimentally resolved complexes (p < 0.01). These findings establish that current RPI prediction benchmarks substantially inflate performance estimates and demonstrate that cross-modal attention architectures yield improved generalization alongside mechanistically interpretable representations.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.