Back

Distilling Mechanistic Models From Multi-Omics Data

Erwin, S.; Fletcher, J. R.; Sweeney, D. C.; Theriot, C. M.; Lanzas, C.

2023-09-06 systems biology
10.1101/2023.09.06.556597 bioRxiv
Show abstract

High-dimensional multi-omics data sets are increasingly accessible and now routinely being generated as part of medical and biological experiments. However, the ability to infer mechanisms of these data remains low due to the abundance of confounding data. The gap between data generation and interpretation highlights the need for strategies to harmonize and distill complex multi-omics data sets into concise, mechanistic descriptions. To this end, a four-step analysis approach for multiomics data is herein demonstrated, comprising: filling missing data and harmonizing data sources, inducing sparsity, developing mechanistic models, and interpretation. This strategy is employed to generate a parsimonious mechanistic model from high-dimensional transcriptomics and metabolomics data collected from a murine model of Clostridioides difficile infection. This approach highlighted the role of the Stickland reactor in the production of toxins during infection, in agreement with recent literature. The methodology present here is demonstrated to be feasible for interpreting multi-omics data sets and it, to the authors knowledge, one of the first reports of a successful implementation of such a strategy.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.