MetaDIA: A Novel Database Reduction Strategy for DIA Human Gut Metaproteomics
Duan, H.; Ning, Z.; Sun, Z.; Guo, T.; Sun, Y.; Figeys, D.
Show abstract
BackgroundMicrobiomes, especially within the gut, are complex and may comprise hundreds of species. The identification of peptides in metaproteomics presents a significant challenge, as it involves matching peptides to mass spectra within an enormous search space for complex and unknown samples. This poses difficulties for both the accuracy and the speed of identification. Specifically, analysis of data-independent acquisition (DIA) datasets has relied on libraries constructed from prior data-dependent acquisition (DDA) results. This approach requires running the samples in DDA mode to construct a library from the identified results, which can then be used for the DIA data. However, this method is resource-intensive, consumes samples, and limits identification to peptides previously identified by DDA. These limitations restrict the application of DIA in metaproteomics research. ResultsWe introduced a novel strategy to reduce the search space by utilizing species abundance and functional abundance information from the microbiome to score each peptide and prioritize those most likely to be detected. Employing this strategy, we have developed and optimized a workflow called MetaDIA for analysis of microbiome DIA data, which operates independently of DDA assistance. Our method demonstrated strong consistency with the traditional DDA-based library approach at both protein and functional levels. ConclusionOur approach successfully created a smaller, yet sufficient database for DIA data search requirements in metaproteomics, showing high consistency with results from the conventional DDA-based library. We believe this method can facilitate the application of DIA in metaproteomics.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A sectioning and database enrichment approach for improved peptide spectrum matching in large, genome-guided protein sequence databases 96%
- Biological Function Assignment Across Taxonomic Levels in Mass-Spectrometry-Based Metaproteomics via a Modified Expectation Maximization Algorithm 96%
- A proteogenomic resource enabling integrated analysis of Listeria genotype-proteotype-phenotype relationships 95%
Similar papers in this journal
- An economic and robust TMT labeling approach for high throughput proteomic and metaproteomic analysis 96%
- Data-Independent Acquisition Mass Spectrometry as a Tool for Metaproteomics: Interlaboratory Comparison Using a Model Microbiome 94%
- Monitoring Functional Post-Translational Modifications Using a Data-Driven Proteome Informatic Pipeline 94%
Similar papers in this journal
- Critical Assessment of MetaProteome Investigation 2 (CAMPI-2): Multi-laboratory assessment of sample processing methods to stabilize fecal microbiome for functional analysis 97%
- Contigs directed gene annotation (ConDiGA) for accurate protein sequence database construction in metaproteomics 97%
- RapidAIM: A culture- and metaproteomics-based Rapid Assay of Individual Microbiome responses to drugs 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.