Phoenix Enhancer: proteomics data mining using clustered spectra
Bai, M.; Qin, C.; Shu, K.; Griss, J.; Perez-Riverol, Y.; Zhu, W.; Hermjakob, H.
Show abstract
MotivationSpectrum clustering has been used to enhance proteomics data analysis: some originally unidentified spectra can potentially be identified and individual peptides can be evaluated to find potential mis-identifications by using clusters of identified spectra. The Phoenix Enhancer provides an infrastructure to analyze tandem mass spectra and the corresponding peptides in the context of previously identified public data. Based on PRIDE Cluster data and a newly developed pipeline, four functionalities are provided: i) evaluate the original peptide identifications in an individual dataset, to find low confidence peptide spectrum matches (PSMs) which could correspond to mis-identifications; ii) provide confidence scores for all originally identified PSMs, to help users evaluate their quality (complementary to getting a global false discovery rate); iii) identify potential new PSMs for originally unidentified spectra; and iv) provide a collection of browsing and visualization tools to analyze and export the results. In addition to the web based service, the code is open-source and easy to re-deploy on local computers using Docker containers. AvailabilityThe service of Phoenix Enhancer is available at http://enhancer.ncpsb.org. All source code is freely available in GitHub (https://github.com/phoenix-cluster/) and can be deployed in the Cloud and HPC architectures. Contactbaimz@cqupt.edu.cn Supplementary informationSupplementary data are available online.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Improved open modification searching via unified spectral search with predicted libraries and enhanced vector representations in ANN-SoLo 98%
- MSnbase, efficient and elegant R-based processing and visualisation of raw mass spectrometry data 97%
- rawR - Direct access to raw mass spectrometry data in R 97%
Similar papers in this journal
- Mistle: bringing spectral library predictions to metaproteomics with an efficient search index 97%
- Alpha-XIC: a deep neural network for scoring the coelution of peak groups improves peptide identification by data-independent acquisition mass spectrometry 97%
- TIMSCONVERT: A workflow to convert trapped ion mobility data to open data formats 96%
Similar papers in this journal
Similar papers in this journal
- Fast alignment of mass spectra in large proteomics datasets, capturing dissimilarities arising from multiple complex modifications of peptides 96%
- Moiety Modeling Framework for Deriving Moiety Abundances from Mass Spectrometry Measured Isotopologues 93%
- MassComp, a lossless compressor for mass spectrometry data 93%
Similar papers in this journal
- promor: a comprehensive R package for label-free proteomics data analysis and predictive modeling 94%
- Enrichment analysis for spatial and single-cell metabolomics accounting for molecular ambiguity 92%
- Covariate balanced allocation of samples to batches to mitigate the impacts of technical variability. 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.