Disentangling covariate effects on single cell-resolved epigenomes with DeepDive
Moeller, A.; Madsen, J.
Show abstract
Understanding the effects of individual biological factors from single cell-resolved epigenomic data is hindered by multicollinearity, particularly in human cohorts. We introduce DeepDive, a novel deep learning framework designed to systematically disentangle known and unknown sources of variation in single-nucleus ATAC-seq data. DeepDive accurately reconstructs chromatin accessibility, outperforms state-of-the-art methods with incomplete covariate information, and robustly recovers true biological signals from even highly entangled covariates, unlocking counter-factual, what-if, analyses. Applying DeepDive to pancreatic islet cells, we perform counter-factual analyses to prioritize covariates associated with a type 2 diabetes-linked beta cell subtype and nominate transcription regulators. DeepDive offers a powerful and unbiased tool for mechanistic discovery in complex human disease cohorts.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Pathway Centric Analysis for single-cell RNA-seq and Spatial Transcriptomics Data with GSDensity 96%
- PACS allows comprehensive dissection of multiple factors governing chromatin accessibility from snATAC-seq data 96%
- NiCo Identifies Extrinsic Drivers of Cell State Modulation by Niche Covariation Analysis 96%
Similar papers in this journal
Similar papers in this journal
- Single-cell DNA methylome and 3D genome atlas of the human subcutaneous adipose tissue 96%
- Benchmarking of deep neural networks for predicting personal gene expression from DNA sequence highlights shortcomings 96%
- ArchR: An integrative and scalable software package for single-cell chromatin accessibility analysis 95%
Similar papers in this journal
- A read count-based method to detect multiplets and their cellular origins from snATAC-seq data 97%
- Smoother: A Unified and Modular Framework for Incorporating Structural Dependency in Spatial Omics Data 96%
- STHD: probabilistic cell typing of single Spots in whole Transcriptome spatial data with High Definition 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.