Biomarker Candidates for Tumors Identified from Deep-Profiled Plasma Stem Predominantly from the Low Abundant Area
Tognetti, M.; Sklodowski, K.; Mueller, S.; Kamber, D.; Muntel, J.; Bruderer, R.; Reiter, L.
Show abstract
The plasma proteome has the potential to enable a holistic analysis of the health state of an individual. However, plasma biomarker discovery is difficult due to its high dynamic range and variability. Here, we present a novel automated analytical approach for deep plasma profiling and applied it to a 180-sample cohort of human plasma from lung, breast, colorectal, pancreatic, and prostate cancer. Using a controlled quantitative experiment, we demonstrate a 257% increase in protein identification and a 263% increase in significantly differentially abundant proteins over neat plasma. In the cohort, we identified 2,732 proteins. Using machine learning, we discovered biomarker candidates such as STAT3 in colorectal cancer and developed models that classify the disease state. For pancreatic cancer, a separation by stage was achieved. Importantly, biomarker candidates came predominantly from the low abundance region, demonstrating the necessity to deeply profile because they would have been missed by shallow profiling.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Automated sample preparation with SP3 for low-input clinical proteomics 96%
- Turnover and replication analysis by isotope labeling (TRAIL) reveals the influence of tissue context on protein and organelle lifetimes 95%
- Causal integration of multi-omics data with prior knowledge to generate mechanistic hypotheses 94%
Similar papers in this journal
- To fly, or not to fly, that is the question: A deep learning model for peptide detectability prediction in mass spectrometry 95%
- Quantitative analysis of non-histone lysine methylation sites and lysine demethylases in breast cancer cell lines 95%
- In vitro Kinase-to-Phosphosite database (iKiP-DB) predicts kinase activity in phosphoproteomic datasets 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.