Extending PROXIMAL to predict degradation pathways of phenolic compounds in the human gut microbiota
Balzerani, F.; Blasco, T.; Perez, S.; Valcarcel, L. V.; Planes, F. J.; Hassoun, S.
Show abstract
Despite significant advances in reconstructing genome-scale metabolic networks, the understanding of cellular metabolism remains incomplete for many organisms. A promising approach for elucidating cellular metabolism is analysing the full scope of enzyme promiscuity, which exploits the capacity of enzymes to bind to non-annotated substrates and generate novel reactions. To guide time-consuming costly experimentation, different computational methods have been proposed for exploring enzyme promiscuity. One relevant algorithm is PROXIMAL, which strongly relies on KEGG to define generic reaction rules and link specific molecular substructures with associated chemical transformations. Here, we present a completely new pipeline, PROXIMAL2, which overcomes the dependency on KEGG data. In addition, PROXIMAL2 introduces two relevant improvements with respect to the former version: i) correct treatment of multi-step reactions and ii) tracking of electric charges in the transformations. We compare PROXIMAL and PROXIMAL2 in recovering annotated products from substrates in KEGG reactions, finding a highly significant improvement in the level of accuracy. We then applied PROXIMAL2 to predict degradation reactions of phenolic compounds in the human gut microbiota. The results were compared to RetroPath RL, a different and relevant enzyme promiscuity method. We found a significant overlap between these two methods but also complementary results, which open new research directions into this relevant question in nutrition.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- dGPredictor: Automated fragmentation method for metabolic reaction free energy prediction and de novo pathway design 94%
- Ranking microbial metabolomic and genomic links in the NPLinker framework using complementary scoring functions 94%
- iPRESTO: automated discovery of biosynthetic sub-clusters linked to specific natural product substructures 94%
Similar papers in this journal
- Benchmark dataset for training machine learning models to predict the pathway involvement of metabolites 94%
- Hierarchical Harmonization of Atom-Resolved Metabolic Re-actions Across Metabolic Databases 94%
- Atom Identifiers Generated by a Graph Coloring Method Enable Compound Harmonization Across Metabolic Databases 94%
Similar papers in this journal
- The ModelSEED Database for the integration of metabolic annotations and the reconstruction, comparison, and analysis of metabolic models for plants, fungi, and microbes 94%
- A Deep Learning Genome-Mining Strategy Improves Biosynthetic Gene Cluster Prediction 93%
- MetaNetX/MNXref - unified namespace for metabolites and biochemical reactions in the context of metabolic models 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.