Semi-Automatic Detection of Errors in Genome-Scale Metabolic Models
Moyer, D. C.; Reimertz, J.; Segre, D.; Fuxman Bass, J. I.
Show abstract
BackgroundGenome-Scale Metabolic Models (GSMMs) are used for numerous tasks requiring computational estimates of metabolic fluxes, from predicting novel drug targets to engineering microbes to produce valuable compounds. A key limiting step in most applications of GSMMs is ensuring their representation of the target organisms metabolism is complete and accurate. Identifying and visualizing errors in GSMMs is complicated by the fact that they contain thousands of densely interconnected reactions. Furthermore, many errors in GSMMs only become apparent when considering pathways of connected reactions collectively, as opposed to examining reactions individually. ResultsWe present Metabolic Accuracy Check and Analysis Workflow (MACAW), a collection of algorithms for detecting errors in GSMMs. The relative frequencies of errors we detect in manually curated GSMMs appear to reflect the different approaches used to curate them. Changing the method used to automatically create a GSMM from a particular organisms genome can have a larger impact on the kinds of errors in the resulting GSMM than using the same method with a different organisms genome. Our algorithms are particularly capable of identifying errors that are only apparent at the pathway level, including loops, and nontrivial cases of dead ends. ConclusionsMACAW is capable of identifying inaccuracies of varying severity in a wide range of GSMMs. Correcting these errors can measurably improve the predictive capacity of a GSMM. The relative prevalence of each type of error we identify in a large collection of GSMMs could help shape future efforts for further automation of error correction and GSMM creation.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Flux-based hierarchical organization of Escherichia coli’s metabolic network 96%
- Quantifying cumulative phenotypic and genomic evidence for procedural generation of metabolic network reconstructions 96%
- Transcriptome-guided parsimonious flux analysis improves predictions with metabolic networks in complex environments 95%
Similar papers in this journal
Similar papers in this journal
- Functional Decomposition of Metabolism allows a system-level quantification of fluxes and protein allocation towards specific metabolic functions 94%
- Data integration across conditions improves turnover number estimates and metabolic predictions 93%
- Projecting genetic associations through gene expression patterns highlights disease etiology and drug mechanisms 93%
Similar papers in this journal
- Automatic reconstruction of metabolic pathways from identified biosynthetic gene clusters 94%
- The effects of model complexity and size on metabolic flux distribution and control. Case study in E. coli. 93%
- SCOUR: A stepwise machine learning framework for predicting metabolite-dependent regulatory interactions 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.