Back

Integrated View of Baseline Protein Expression in Human Tissues using public Data Independent Acquisition datasets

Prakash, A.; Collins, A.; Vilmovsky, L.; Fexova, S.; Vizcaino, J. A.; Jones, A. R.

2024-09-19 bioinformatics
10.1101/2024.09.16.613191 bioRxiv
Show abstract

The PRIDE database is the largest public data repository of mass spectrometry-based proteomics data and currently stores more than 40,000 datasets covering a wide range of organisms, experimental techniques and biological conditions. During the past few years, PRIDE has seen a significant increase in the amount of submitted Data-Independent Acquisition (DIA) proteomics datasets. This provides an excellent opportunity for large scale data reanalysis and reuse. We have reanalysed 15 public label-free DIA datasets across various healthy human tissues, to provide a state-of-the-art view of the human proteome in baseline conditions (without any perturbations). We computed baseline protein abundances and compared them across various tissues, samples and datasets. Our second aim was to compare protein abundances obtained here from the results of previous analyses using human baseline Data-Dependent Acquisition (DDA) datasets. We observed a good correlation across some tissues, especially in liver and colon but weak correlations were found in others, such as lung and pancreas. The reanalysed results including protein abundance values and curated metadata are made available to view and download from the resource Expression Atlas. For TOC Only O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=126 SRC="FIGDIR/small/613191v2_ufig1.gif" ALT="Figure 1"> View larger version (26K): org.highwire.dtl.DTLVardef@1414dforg.highwire.dtl.DTLVardef@66606aorg.highwire.dtl.DTLVardef@143f862org.highwire.dtl.DTLVardef@1682406_HPS_FORMAT_FIGEXP M_FIG C_FIG

Matching journals

The top 1 journal accounts for 50% of the predicted probability mass.