MuSHIN: A multi-way SMILES-based hypergraph inference network for metabolic model reconstruction
Zhao, Y.; Chen, Y.; Yu, Y.; Liu, X.; Du, J.; Wen, J.; Liao, C.; Sun, Q.; Wang, R.; Chen, C.
Show abstract
Genome-scale metabolic models (GEMs) are indispensable tools for probing cellular metabolism, enabling predictions of metabolic fluxes, guiding strain optimization, and advancing biomedical research. However, their predictive capacity is often compromised by incomplete reaction networks, stemming from gaps in biochemical knowledge, annotation inaccuracies, and insufficient experimental validations. Here we present MuSHIN (Multi-way SMILES-based Hypergraph Interface Network), a novel deep hypergraph learning method that integrates network topology with biochemical domain knowledge to predict missing reactions in GEMs. Evaluated on 926 high- and intermediate-quality GEMs with artificially removed reactions, MuSHIN significantly outperforms state-of-the-art methods, achieving up to a 17% improvement across multiple metrics and maintaining robust recovery even under severe network sparsity. Furthermore, MuSHIN substantially enhances phenotypic predictions in 24 draft GEMs associated with fermentation by resolving critical metabolic gaps, as validated against experimental measurements. Together, these findings highlight MuSHINs potential to advance GEM reconstruction and accelerate discoveries in systems biology, metabolic engineering, and precision medicine.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Engineering of highly active and diverse nuclease enzymes by combining machine learning and ultra-high-throughput screening 95%
- scTrace+: enhance the cell fate inference by integrating the lineage-tracing and multi-faceted transcriptomic similarity information 94%
- Integration of multi-modal measurements identifies critical mechanisms of tuberculosis drug action 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.