Boosting metaproteomics identification rates and taxonomic specificity with MS2Rescore
Van Den Bossche, T.; Declercq, A.; Gabriels, R.; Holstein, T.; Mesuere, B.; Muth, T.; Verschaffelt, P.; Martens, L.
Show abstract
BackgroundMetaproteomics, the study of the collective proteome within microbial ecosystems, has gained increasing interest over the past decade. However, peptide identification rates in metaproteomics remain low compared to single-species proteomics. A key challenge is the identification sensitivity of current identification algorithms, which were primarily designed for single-species analyses. Addressing this, we evaluated the machine learning-driven MS{superscript 2}Rescore post-processing tool on multiple metaproteomics datasets from diverse microbial environments and benchmark studies. ResultsWe demonstrate that machine learning-driven rescoring outperforms traditional metaproteomics identification workflows. It significantly increases peptide identification rates compared to Sage, which itself already implements basic rescoring. Moreover, it enables lowering the false discovery rate (FDR) to 0.1% with minimal to no sensitivity loss, a substantial improvement over the 1% or 5% FDR thresholds commonly used in metaproteomics, in turn leading to greater confidence in downstream taxonomic annotation. ConclusionsOur findings show that MS{superscript 2}Rescore substantially improves peptide identification sensitivity as well as specificity in metaproteomics, and delivers improved confidence in taxonomic annotation. This advancement results in a more reliable downstream taxonomic analysis, reinforcing the potential of machine learning-based rescoring in metaproteomics research.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A sectioning and database enrichment approach for improved peptide spectrum matching in large, genome-guided protein sequence databases 98%
- The use of hybrid data-dependent and -independent acquisition spectral libraries empower dual-proteome profiling 97%
- Fast and memory efficient searching of large-scale mass spectrometry data using Tide 96%
Similar papers in this journal
- An economic and robust TMT labeling approach for high throughput proteomic and metaproteomic analysis 97%
- Leveraging immonium ions for identifying and targeting acyl-lysine modifications in proteomic datasets 96%
- Data-Independent Acquisition Mass Spectrometry as a Tool for Metaproteomics: Interlaboratory Comparison Using a Model Microbiome 96%
Similar papers in this journal
- Comparative Performance of Scribe and Database Search Engines in Metaproteomic Profiling of a Ground-Truth Microbiome Dataset 97%
- AA_stat: intelligent profiling of in vivo and in vitro modifications from open search results 96%
- ReCom: A semi-supervised approach to ultra-tolerant database search for improved identification of modified peptides 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.