Back

Visualization and normalization of drift effect across batches in metabolome-wide association studies

Bararpour, N.; Gilardi, F.; Carmeli, C.; Sidibe, J.; Ivanisevic, J.; Caputo, T.; Augsburger, M.; Grabherr, S.; Desvergne, B.; Guex, N.; Bochud, M.; Thomas, A.

2020-01-22 bioinformatics
10.1101/2020.01.22.914051 bioRxiv
Show abstract

As a powerful phenotyping technology, metabolomics provides new opportunities in biomarker discovery through metabolome-wide association studies (MWAS) and identification of metabolites having regulatory effect in various biological processes. While MS-based metabolomics assays are endowed with high-throughput and sensitivity, large-scale MWAS are doomed to long-term data acquisition generating an overtime-analytical signal drift that can hinder the uncovering of true biologically relevant changes. We developed "dbnorm", a package in R environment, which allows visualization and removal of signal heterogeneity from large metabolomics datasets. "dbnorm" integrates advanced statistical tools to inspect dataset structure, at both macroscopic (sample batch) and microscopic (metabolic features) scales. To compare model performance on data correction, "dbnorm" assigns a score, which allows the straightforward identification of the best fitting model for each dataset. Herein, we show how "dbnorm" efficiently removes signal drift among batches to capture the true biological heterogeneity of data in two large-scale metabolomics studies.

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.