Genome-Wide Uncertainty-Moderated Extraction of Signal Annotations from Multi-Sample Functional Genomics Data
Hamilton, N. H.; McMichael, B. D.; Love, M. I.; Furey, T. S.
Show abstract
We present Consenrich, a simple but principled technique for genome-wide estimation of signals hidden in noisy multi-sample sequencing-based functional genomics datasets. Consenrich appeals to a sequential prediction-correction framework and models both the spatial dependencies between proximal loci and regional, sample-specific noise processes that corrupt sequencing data. Experiments reveal distinct improvement compared to benchmarks in a series of challenging estimation problems, where noisy functional genomics data samples must be reconciled. We further highlight the immediate practical appeal of this refined signal extraction for differential analyses between disease conditions and identification of functionally enriched genomic regions. A complete implementation of Consenrich is hosted at https://github.com/nolan-h-hamilton/Consenrich.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- GoM DE: interpreting structure in sequence count data with differential expression analysis allowing for grades of membership 95%
- Robustness and applicability of functional genomics tools on scRNA-seq data 95%
- Simultaneous smoothing and detection of topological units of genome organization from sparse chromatin contact count matrices with matrix factorization 95%
Similar papers in this journal
Similar papers in this journal
- Characterizing the properties of bisulfite sequencing data: maximizing power and sensitivity to identify between-group differences in DNA methylation 94%
- Copy number normalization distinguishes differential signals driven by copy number differences in ATAC-seq and ChIP-seq 94%
- Neural network modeling of differential binding between wild-type and mutant CTCF reveals putative binding preferences for zinc fingers 1-2 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.