Back

Deep plasma proteomics with data-independent acquisition: A fastlane towards biomarkers identification.

Ward, B.; Pyr dit Ruys, S.; Balligand, J.-L.; Belkhir, L.; Cani, P. D.; Collet, J.-F.; De Greef, J.; Gatto, L.; Haufroid, V.; Jodogne, S.; Kabamba, B.; Lingurski, M.; Yombi, J. C.; Vertommen, D.; Elens, L.

2024-02-26 systems biology
10.1101/2024.02.23.581160 bioRxiv
Show abstract

Plasma proteomic is a precious tool in human disease research, but requires extensive sample preparation in order to perform in-depth analysis and biomarker discovery using traditional Data-Dependent Acquisition (DDA). Here, we highlight the efficacy of combining moderate plasma prefractionation and Data-Independent Acquisition (DIA) to significantly improve proteome coverage and depth, while remaining cost- and time-efficient. Using human plasma collected from a 20-patient COVID-19 cohort, our method utilises commonly available solutions for depletion, sample preparation, and fractionation, followed by 3 LC-MS/MS injections for a 360-minutes DIA run time. DIA-NN software was then used for precursor identification, and the QFeatures R package was used for protein aggregation. We detect 1,321 proteins on average per patient, and 2,031 unique proteins across the cohort. Filtering precursors present in under 25% of patients, we still detect 1,230 average proteins and 1,590 unique proteins, indicating robust protein identification. Differential analysis further demonstrates the applicability of this method for plasma proteomic research and clinical biomarker identification. In summary, this study introduces a streamlined, cost- and time-effective approach to deep plasma proteome analysis, expanding its utility beyond classical research environments and enabling larger-scale multi-omics investigations in clinical settings.

Matching journals

The top 1 journal accounts for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.