Investigation of normalization procedures for transcriptome profiles of compounds oriented toward practical study design
Mizuno, T.; Kusuhara, H.
Show abstract
The transcriptome profile is a representative phenotype-based descriptor of compounds, widely acknowledged for its ability to effectively capture compound effects. However, the presence of batch differences is inevitable. Despite the existence of sophisticated statistical methods, many of them presume a substantial sample size. How should we design a transcriptome analysis to obtain robust compound profiles, particularly in the context of small datasets frequently encountered in practical scenarios? This study addresses this question by investigating the normalization procedures for transcriptome profiles, focusing on the baseline distribution employed in deriving biological responses as profiles. Firstly, we investigated two large GeneChip datasets, comparing the impact of different normalization procedures. Through an evaluation of the similarity between response profiles of biological replicates within each dataset and the similarity between response profiles of the same compound across datasets, we revealed that the baseline distribution defined by all samples within each batch under batch-corrected condition is a good choice for large datasets. Subsequently, we conducted a simulation to explore the influence of the number of control samples on the robustness of response profiles across datasets. The results offer insights into determining the suitable quantity of control samples for diminutive datasets. It is crucial to acknowledge that these conclusions stem from constrained datasets. Nevertheless, we believe that this study enhances our understanding of how to effectively leverage transcriptome profiles of compounds and promotes the accumulation of essential knowledge for the practical application of such profiles.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Predicting biological pathways of chemical compounds with a profile-inspired aproach 94%
- Predicting the pathway involvement of metabolites annotated in the MetaCyc knowledgebase 93%
- Quantitative Structure-Mutation-Activity Relationship Tests (QSMART) Model for Protein Kinase Inhibitor Response Prediction 93%
Similar papers in this journal
- Tensor decomposition- and principal component analysis-based unsupervised feature extraction to select more reasonable differentially expressed genes: Optimization of standard deviation versus state-of-art methods 94%
- Evaluation of Connectivity Map shows limited reproducibility in drug repositioning 93%
- A new strategy for identifying mechanisms of drug-drug interaction using transcriptome analysis: Compound Kushen injection as a proof of principle 93%
Similar papers in this journal
- Benchmark dataset for training machine learning models to predict the pathway involvement of metabolites 94%
- Atom Identifiers Generated by a Graph Coloring Method Enable Compound Harmonization Across Metabolic Databases 94%
- Hierarchical Harmonization of Atom-Resolved Metabolic Re-actions Across Metabolic Databases 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.