The Peptonizer2000: bringing confidence to metaproteomics
Holstein, T.; Verschaffelt, P.; Van de Vyver, S.; Van den Bossche, T.; Mesuere, B.; Martens, L.; Muth, T.
Show abstract
Metaproteomics, the large-scale study of proteins from microbial communities, faces challenges in identifying species due to similarities in protein sequences across different organisms. Current methods often rely on simple counting of matches between proteins and taxa, which can lead to low accuracy. We introduce the Peptonizer2000, a new tool that uses advanced modeling to provide more precise taxonomic identifications along with confidence scores. It combines peptide scores from any proteomic search engine with peptide-to-taxon links from the Unipept database. By applying statistical models, the Peptonizer2000 improves taxonomic resolution and delivers more reliable results. We validate its performance using publicly available datasets, demonstrating its ability to produce high-confidence identifications. Our results suggest that the Peptonizer2000 improves the specificity and confidence of taxonomic assignments in metaproteomics, providing a valuable resource for the study of complex microbial communities.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- Biological Function Assignment Across Taxonomic Levels in Mass-Spectrometry-Based Metaproteomics via a Modified Expectation Maximization Algorithm 98%
- A sectioning and database enrichment approach for improved peptide spectrum matching in large, genome-guided protein sequence databases 98%
- TaxIt: An iterative and automated computational pipeline for untargeted strain-level identification using MS/MS spectra from pathogenic samples 96%
Similar papers in this journal
- Critical Assessment of Metaproteome Investigation (CAMPI): a Multi-Lab Comparison of Established Workflows 97%
- DeepRTAlign: toward accurate retention time alignment for large cohort mass spectrometry data analysis 95%
- MSFragger-DDA+ Enhances Peptide Identification Sensitivity with Full Isolation Window Search 94%
Similar papers in this journal
- Common data models to streamline metabolomics processing and annotation, and implementation in a Python pipeline 95%
- Multienzyme deep learning models improve peptide de novo sequencing by mass spectrometry proteomics 95%
- Spec2Vec: Improved mass spectral similarity scoring through learning of structural relationships 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.