Ratio-based quantitative multiomics profiling using universal reference materials empowers data integration
Zheng, Y.; Liu, Y.; Yang, J.; Dong, L.; Zhang, R.; Tian, S.; Yu, Y.; Ren, L.; Hou, W.; Zhu, F.; Mai, Y.; Han, J.; Zhang, L.; Jiang, H.; Lin, L.; Lou, J.; Li, R.; Lin, J.; Liu, H.; Kong, Z.; Wang, D.; Dai, F.; Bao, D.; Cao, Z.; Chen, Q.; Chen, Q.; Chen, X.; Gao, Y.; Jiang, H.; Li, B.; Li, B.; Li, J.; Liu, R.; Qing, T.; Shang, E.; Shang, J.; Sun, S.; Wang, H.; Wang, X.; Zhang, N.; Zhang, P.; Zhang, R.; Zhu, S.; Scherer, A.; Wang, J.; Wang, J.; Xu, J.; Hong, H.; Xiao, W.; Liang, X.; Jin, L.; The Quartet Project Team, ; Tong, W.; Ding, C.; Li, J.; Fang, X.; Shi, L.
Show abstract
Multiomics profiling is a powerful tool to characterize the same samples with complementary features orchestrating the genome, epigenome, transcriptome, proteome, and metabolome. However, the lack of ground truth hampers the objective assessment of and subsequent choice from a plethora of measurement and computational methods aiming to integrate diverse and often enigmatically incomparable omics datasets. Here we establish and characterize the first suites of publicly available multiomics reference materials of matched DNA, RNA, proteins, and metabolites derived from immortalized cell lines from a family quartet of parents and monozygotic twin daughters, providing built-in truth defined by family relationship and the central dogma. We demonstrate that the "ratio"-based omics profiling data, i.e., by scaling the absolute feature values of a study sample relative to those of a concurrently measured universal reference sample, were inherently much more reproducible and comparable across batches, labs, platforms, and omics types, thus empower the horizontal (within-omics) and vertical (cross-omics) data integration in multiomics studies. Our study identifies "absolute" feature quantitation as the root cause of irreproducibility in multiomics measurement and data integration, and urges a paradigm shift from "absolute" to "ratio"-based multiomics profiling with universal reference materials.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- LEOPARD: missing view completion for multi-timepoints omics data via representation disentanglement and temporal knowledge transfer 96%
- TidyMass2: Advancing LC-MS Untargeted Metabolomics Through Metabolite Origin Inference and Metabolic Feature-based Functional Module Analysis 95%
- TrimNN: Characterizing cellular community motifs for studying multicellular topological organization in complex tissues 95%
Similar papers in this journal
- Prioritized single-cell proteomics reveals molecular and functional polarization across primary macrophages 95%
- High-Parameter Spatial Multi-Omics through Histology-Anchored Integration 94%
- ISSAAC-seq enables sensitive and flexible multimodal profiling of chromatin accessibility and gene expression in single cells 94%
Similar papers in this journal
- Turnover and replication analysis by isotope labeling (TRAIL) reveals the influence of tissue context on protein and organelle lifetimes 93%
- PIFiA: Self-supervised Approach for Protein Functional Annotation from Single-Cell Imaging Data 93%
- hu.MAP3.0: Atlas of human protein complexes by integration of > 25,000 proteomic experiments 93%
Similar papers in this journal
- Correcting batch effects in large-scale multiomic studies using a reference-material-based ratio method 98%
- stGCL: A versatile cross-modality fusion method based on multi-modal graph contrastive learning for spatial transcriptomics 94%
- DEMINERS enables clinical metagenomics and comparative transcriptomic analysis by increasing throughput and accuracy of nanopore direct RNA sequencing 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.