Back

Distortion Discovery: A Framework to Model, Spot and Explain Tumor Heterogeneity and Mitigate its Negative Impact on Cancer Risk Assessment

Elmansy, D. F.

2021-04-29 bioinformatics
10.1101/2021.04.28.441787 bioRxiv
Show abstract

In a complex system of inter-genome interactions, false negatives remain an overwhelming problem when using omics data for disease risk prediction. This is especially clear when dealing with complex diseases like cancer in which the infiltration of stromal and immune cells into the tumor tissue can affect the degree of its tumor purity and hence its cancer signal. Previous work was done to estimate the degree of cancer purity in a tissue. In this work, we introduce a data and biomarker selection independent, information theoretic, approach to tackle this problem. We model distortion as a source of false negatives and introduce a mechanism to detect and remove its impact on the accuracy of disease risk prediction.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.