ProteoParc: A tool to generate protein reference databases for ancient and non-model organisms
Carrillo-Martin, G.; Krueger, J.; Marques-Bonet, T.; Lizano, E.
Show abstract
Over the last few years, the increasing interest in analysing the proteome of extinct and non-model organisms has generated a new field of research expanding the scope of proteomics. The lack of curated databases and/or molecular data from these organisms forces researchers to manually search in different public repositories for related protein sequences, either for MS/MS peptide identification or ZooMS marker annotation. This can lead to format incongruences and hinder reproducibility between studies. To address this issue, we introduce ProteoParc, a user-friendly software that generates reference databases by systematically downloading and processing protein sequences from the most widely used public repositories. The pipelines output is a non-redundant protein database, formatted to be interpreted by typical peptide identification software. Moreover, the user can adjust the database dimension and composition by applying different criteria to include only a certain number of genes or species. Thus, ProteoParc is an easy and fast, custom-made bioinformatic tool useful for future paleoproteomics analysis in ancient samples related to understudied organisms.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- VIQoR: a web service for Visually supervised protein Inference and protein Quantification 95%
- Quickomics: exploring omics data in an intuitive, interactive and informative manner 95%
- LimROTS: A Hybrid Method Integrating Empirical Bayes and Reproducibility-Optimized Statistics for Robust Differential Expression Analysis 94%
Similar papers in this journal
Similar papers in this journal
- Parchment Glutamine Index (PQI): A novel method to estimate glutamine deamidation levels in parchment collagen obtained from low-quality MALDI-TOF data 93%
- Re-annotation of SARS-CoV-2 proteins using an HHpred-based approach opens new opportunities for a better understanding of this virus 92%
- Palaeoproteomic identification of a whale bone tool from Bronze Age Heiloo, the Netherlands 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.