Predicting strain-specific metabolic capabilities in the Genus Pseudomonas using a Flux-to-AI approach unravels hidden cell envelope properties.
Focil, C.; Dalldorf, C.; Martinez, D.; Zepeda, A.; Zuniga, C.
Show abstract
Bacteria from the Pseudomonas genus are omnipresent in air, soil, and water. They have been widely studied for their broad metabolic versatility and presence in the epidemiological chain, bioproduction, bioremediation, and disease processes. Each year, more genomic sequences are reported in databases and repositories. However, the relationship between the genomic variability of Pseudomonas strains and the diversity in their metabolic capabilities under environmentally relevant phenotypes remains unknown. Additionally, predictive tools for the analysis of different strains in a systematic framework are limited. Here, we reconstructed genome-scale metabolic models (GEMs) of 44 Pseudomonas strains from various environments and investigated their capabilities to metabolize different carbon sources and metabolic intermediaries. By systematically testing the substrate utilization of the models, we demonstrate how GEM-predicted capabilities can differentiate between strains and that high metabolic versatility is associated with the ability of the strains to remove toxic compounds while maintaining core functionalities. Hundreds of model simulations were used as input for a classification schema that uses machine learning algorithms, resulting in the identification of metabolic capabilities that better differentiate between species. Interestingly, transcription and expression models validated these findings by showing how pathways in which those metabolites are members change their proteome allocation across strains (e.g. phenylalanine metabolism as well as in the carbohydrate metabolism). Author summaryA central challenge in understanding how Pseudomonas strains can play symbiotic, competitive, and pathogenic roles depending on their environment can be potentially addressed with advanced computational tools that combine systems biology and machine learning approaches. Remarkable progress in understanding relationships between metabolism and phenotypes of Pseudomonas has been achieved through the collection of multi-omics datasets (e.g. genomic, transcriptomic, proteomics, fluxomics, etc.) from different settings such as air, clinical, soil, and water. However, maximum utilization of those tools has been limited by the lack of computational tools that can capture the metabolism at genome-scale. In this work, we are making available source genome-scale metabolic models for the most studied Pseudomonas strains and generated flux predictions across 44 strains under 425 different growth conditions by changing the carbon source in turn. Predicted growth phenotypes were used as input for machine learning to identify critical metabolic activities that change among strains. Computational resources leveraged here will enable deeper understanding of the commonalities and differences triggering the metabolic capabilities of Pseudomonas strains.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Genome-Scale reconstruction of Paenarthrobacter aurescens TC1 metabolic model towards the study of atrazine bioremediation 97%
- Deciphering the metabolic capabilities of Bifidobacteria using genome-scale metabolic models 96%
- A Penicillium rubens platform strain for secondary metabolite production 95%
Similar papers in this journal
- An updated genome-scale metabolic network reconstruction of Pseudomonas aeruginosa PA14 to characterize mucin-driven shifts in bacterial metabolism 97%
- Mechanistic insights into bacterial metabolic reprogramming from omics-integrated genome-scale models 96%
- Genome-scale metabolic modeling reveals increased reliance on valine catabolism in clinical isolates of Klebsiella pneumoniae 94%
Similar papers in this journal
- Panera: A novel framework for surmounting uncertainty in microbial community modelling using Pan-genera metabolic models 95%
- Aerobicity stimulon in Escherichia coli revealed using multi-scale computational systems biology of adapted respiratory variants 94%
- Small proteins from prokaryotes in marine water column at full ocean depth 93%
Similar papers in this journal
- Marine picoplankton metagenomes from eleven vertical profiles obtained by the Malaspina Expedition in the tropical and subtropical oceans 94%
- Highly accurate long-read HiFi sequencing data for five complex genomes 93%
- CF-Seq, An Accessible Web Application for Rapid Re-Analysis of Cystic Fibrosis Pathogen RNA Sequencing Studies 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.