Analysis of model organism viability through an interspecies pathway comparison pipeline using the dynamic impact approach
Nguyen, A. D.; Bionaz, M.
Show abstract
BackgroundComputational biologists investigate gene expression time-series data using estimation, clustering, alignment, and enrichment methods to make biological sense of the data and provide compelling visualization. While there is an abundance of microarray and RNA-seq data available, interpreting the data while capturing the dynamism of a time-course experiment remains a difficult challenge. Advancements in RNA-seq technologies have allowed us to collect extensive profiles of diverse developmental processes but also requires additional methods for analysis and data integration to capture the increased dynamism. An approach that can both capture the dynamism and direction of change in a time-course experiment in a holistic manner and simultaneously identify which biological pathways are significantly altered is necessary for the interpretation of systems biology data. In addition, there is a need for a method to evaluate the viability of model organisms across different treatments and conditions. By comparing effects of a specific treatment (e.g., a drug) on the target pathway between multiple species and determining pathways with a similar response to biological cues between organisms, we can determine the best animal model for that treatment for future studies. MethodsHere, we present Dynamic Impact Approach with Normalization (DIA-norm), a dynamic pathway analysis tool for the analysis of time-course data without unsupervised dimensionality reduction. We analyzed five datasets of mesenchymal stem cells retrieved from the Gene Expression Omnibus data repository (3 human, 1 mouse cell line, 1 pig) which were differentiated in vitro towards adipogenesis. In the first step, DIA-norm calculated an impact and flux score for each biological term using p-value and fold change. In the second step, these scores were normalized and interpolated using cubic spline. Cross-correlation was then performed between all the data sets with r[≥]0.6 as a benchmark for high correlation as r = 0.7 is the limit of experimental reproducibility. ResultsDIA-norm predicted that the pig was a better model for humans than a mouse for the study of adipogenesis. The pig model had a higher number of correlating pathways with humans (64.5 to 30.5) and higher average correlation (r = 0.51 vs r = 0.46) as compared to mouse model vs human. While not a definitive conclusion, the results are in accordance with prior phylogenetic and disease studies in which pigs are a good model for studying humans, specifically regarding obesity. In addition, DIA-norm identified a larger number of biologically important pathways (approximately 2x number of pathways) versus a comparable enrichment analysis tool, DAVID. DIA-norm also identified some possible pathways of interests for adipogenesis, namely, nitrogen metabolism (r = 0.86), where there is little to no existing literature. ConclusionDIA-norm captured 80+% of biological important pathways and achieved high pathway correlation between species for the vast majority of important adipogenesis pathways. DIA-norm can be used for both time-series pathway analysis and the determination of a model organism. Our findings indicate that DIA-norm can be used to study the effect of any treatment, including drugs, on specific pathways between multiple species to determine the best animal model for that treatment for future studies. The reliability of DIA-norm to provide biological insights compared to enrichment approach tools has been demonstrated in the selected transcriptomic studies by identifying a higher number of total and biologically relevant pathways. DIA-norms final advantage was its easily interpretable graphical outputs that aid in visualizing dynamic changes in expression.
Matching journals
The top 13 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Single-cell transcriptome analysis of embryonic and adult endothelial cells allows to rank the hemogenic potential of post-natal endothelium 93%
- Oxygen level alters energy metabolism in bovine preimplantation embryos 93%
- Identification of miRNA signatures for kidney renal clear cell carcinoma using the tensor-decomposition method 93%
Similar papers in this journal
- Unsupervised tensor decomposition-based method to extract candidate transcription factors as histone modification bookmarks in post-mitotic transcriptional reactivation 93%
- Reconstruction of a generic genome-scale metabolic network for chicken: investigating network connectivity and finding potential biomarkers 93%
- Two-Step Mixed Model Approach to Analyzing Differential Alternative RNA Splicing 93%
Similar papers in this journal
- Tensor decomposition-Based Unsupervised Feature Extraction Applied to Single-Cell Gene Expression Analysis 93%
- Analysis of Pan-Omics Data in Human Interactome Network (APODHIN) 92%
- Evolutionary Perspective And Expression Analysis Of Intronless Genes Highlight The Conservation On Their Regulatory Role 92%
Similar papers in this journal
- Evolutionary Inference Predicts Novel ACE2 Protein Interactions Relevant to COVID-19 Pathologies 93%
- A comprehensive algorithmic dissection yields biomarker discovery and insights into the discrete stage-wise progression of colorectal cancer 91%
- Comparison of rule- and ordinary differential equation-based dynamic model of DARPP-32 signalling network 91%
Similar papers in this journal
- Integrated Analysis of Tissue-specific Gene Expression in Diabetes by Tensor Decomposition Can Identify Possible Associated Diseases. 93%
- Small RNA Sequencing Reveals a Distinct MicroRNA Signature between Glucocorticoid Responder and Glucocorticoid Non-responder Primary Human Trabecular Meshwork Cells after Dexamethasone Treatment 93%
- Identification of differentially expressed genes in the longissimus dorsi muscle of Luchuan and Duroc pigs by transcriptome sequencing 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.