Retention time and fragmentation predictors increase confidence in variant peptide identification
Skiadopoulou, D.; Vasicek, J.; Kuznetsova, K.; Kall, L.; Vaudel, M.
Show abstract
Precision medicine focuses on adapting care to the individual profile of patients, e.g. accounting for their unique genetic makeup. Being able to account for the effect of genetic variation on the proteome holds great promises towards this goal. However, identifying the protein products of genetic variation using mass spectrometry has proven very challenging. Here we show that the identification of variant peptides can be improved by the integration of retention time and fragmentation predictors into a unified proteogenomic pipeline. By combining these intrinsic peptide characteristics using the search-engine post-processor Percolator, we demonstrate improved discrimination power between correct and incorrect peptide-spectrum matches. Our results demonstrate that the drop in performance that is induced when expanding a protein sequence database can be compensated, and hence enabling efficient identification of genetic variation products in proteomics data. We anticipate that this enhancement of proteogenomic pipelines can provide a more refined picture of the unique proteome of patients, and thereby contribute to improving patient care.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- The use of hybrid data-dependent and -independent acquisition spectral libraries empower dual-proteome profiling 97%
- The Crux toolkit for analysis of bottom-up tandem mass spectrometry proteomics data 97%
- A machine learning strategy that leverages large datasets to boost statistical power in small-scale experiments 96%
Similar papers in this journal
- AA_stat: intelligent profiling of in vivo and in vitro modifications from open search results 97%
- A systematic evaluation of yeast sample preparation protocols for spectral identifications, proteome coverage and post-isolation modifications 97%
- Decoding the Impact of Neighboring Amino Acid on MS Intensity Output through Deep Learning 97%
Similar papers in this journal
- MMS2plot: an R package for visualizing multiple MS/MS spectra for groups of modified and non-modified peptides 96%
- Benchmarking accuracy and precision of intensity-based absolute quantification of protein abundances in Saccharomyces cerevisiae 96%
- Leveraging immonium ions for identifying and targeting acyl-lysine modifications in proteomic datasets 96%
Similar papers in this journal
- Assessing the role of trypsin in quantitative plasma- and single-cell proteomics towards clinical application 97%
- A Full Window Data Independent Acquisition Method for DeeperTop-down Proteomics 97%
- Optimization of data-independent acquisition using predicted libraries for deep and accurate proteome profiling 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.