Prosit-XL: enhanced cross-linked peptide identification by accurate fragment intensity prediction to study protein-protein interactions and protein structures
Kalhor, M.; Saylan, C. C.; Picciani, M.; Fischer, L.; Schimweg, F.; Lapin, J.; Rappsilber, J.; Wilhelm, M.
Show abstract
It has been shown that integrating peptide property predictions such as fragment intensity into the scoring process of peptide spectrum match can greatly increase the number of confidently identified peptides compared to using traditional scoring methods. Here, we introduce Prosit-XL, a robust and accurate fragment intensity predictor covering the cleavable (DSSO/DSBU) and non-cleavable cross-linkers (DSS/BS3), achieving high accuracy on various holdout sets with consistent performance on external datasets without fine-tuning. Due to the complex nature of false positives in XL-MS, a novel approach to data-driven rescoring was developed that benefits from Prosit-XLs predictions while limiting the overestimation of the false discovery rate (FDR). We first evaluated this approach using two ground truth datasets that demonstrate the accurate and precise FDR estimation. Second, we applied Prosit-XL on a proteome-scale dataset, demonstrating an up to [~]3.4-fold improvement in PPI discovery compared to classic approaches. Finally, Prosit-XL was used to increase the coverage and depth of a spatially resolved interactome map of intact human cytomegalovirus virions, leading to the discovery of previously unobserved interactions between human and cytomegalovirus proteins.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- AlphaPeptDeep: A modular deep learning framework to predict peptide properties for proteomics 98%
- Sequence-to-sequence translation from mass spectra to peptides with a transformer model 97%
- Retention Time Prediction Using Neural Networks Increases Identifications in Crosslinking Mass Spectrometry 97%
Similar papers in this journal
- To fly, or not to fly, that is the question: A deep learning model for peptide detectability prediction in mass spectrometry 97%
- Comparative analysis of chemical cross-linking mass spectrometry data indicates that protein STY residues rarely react with N-hydroxysuccinimide ester cross-linkers 97%
- Searching for Sulfotyrosines (sY) in a HA(pY)STACK 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.