Teasing out Missing Reactions in Genome-scale Metabolic Networks through Deep Learning
Chen, C.; Liao, C.; Liu, Y.-Y.
Show abstract
GEnome-scale Metabolic models (GEMs) are powerful tools to predict cellular metabolism and physiological states in living organisms. However, due to our imperfect knowledge of metabolic processes, even highly curated GEMs have knowledge gaps (e.g., missing reactions). Existing gap-filling methods typically require phenotypic data as input to tease out missing reactions. We still lack a computational method for rapid and accurate gap-filling of metabolic networks before experimental data is available. Here we present a deep learning-based method -- CHEbyshev Spectral HyperlInk pREdictor (CHESHIRE) -- to predict missing reactions in GEMs purely from metabolic network topology. We demonstrate that CHESHIRE outperforms other topology-based methods in predicting artificially removed reactions over 926 high- and intermediate-quality GEMs. Furthermore, CHESHIRE is able to improve the phenotypic predictions of 49 draft GEMs for fermentation products and amino acids secretions. Both types of validation suggest that CHESHIRE is a powerful tool for GEM curation to reveal unknown links between reactions and observed metabolic phenotypes.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- SemiBin2: self-supervised contrastive learning leads to better MAGs for short- and long-read sequencing 94%
- Statistical batch-aware embedded integration, dimension reduction and alignment for spatial transcriptomics 93%
- SMILE: Mutual Information Learning for Integration of Single Cell Omics Data 92%
Similar papers in this journal
- High-precision cell-type mapping and annotation of single-cell spatial transcriptomics with STAMapper 94%
- Characterizing Spatially Continuous Variations in Tissue Microenvironment through Niche Trajectory Analysis 94%
- Multi-omics analysis reveals the molecular response to heat stress in a "red tide" dinoflagellate 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.