Back

UniScore, a unified and universal measure for peptide identification by multiple search engines

Tabata, T.; Yoshizawa, A. C.; Ogata, K.; Chang, C.-H.; Araki, N.; Sugiyama, N.; Ishihama, Y.

2024-10-13 bioinformatics
10.1101/2024.10.09.617445 bioRxiv
Show abstract

We propose UniScore as a metric for integrating and standardizing the outputs of multiple search engines in the analysis of data-dependent acquisition (DDA) data from LC/MS/MS-based bottom-up proteomics. UniScore is calculated from the annotation information attached to the product ions alone by matching the amino acid sequences of candidate peptides suggested by the search engine with the product ion spectrum. The acceptance criteria are controlled independently of the score values by using the false discovery rate based on the target-decoy approach. Compared to other rescoring methods that use deep learning-based spectral prediction, larger amounts of data can be processed using minimal computing resources. When applied to large-scale global proteome data and phosphoproteome data, the UniScore approach outperformed each of the conventional single search engines examined (Comet, X! Tandem, Mascot and MaxQuant). Furthermore, UniScore could also be directly applied to peptide matching in chimeric spectra without any additional filters.

Matching journals

The top 1 journal accounts for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.