Addressing technical variations in ATAC-seq data and improving motif accessibility analyses
Wang, J.; Sonder, E.; Domcke, S.; Robinson, M. D.; Germain, P.-L.
Show abstract
Tagmentation-based methods such as ATAC-seq and Cut&Tag have provided easy ways to profile the epigenome in low-input samples and even single cells. In this contribution, we discuss forms of bias (i.e. technical variations) in tagmentation-based data, in particular ATAC-seq, and introduce three R/bioconductor packages to facilitate bulk and single-cell epigenomic data analysis, with a special focus on motif accessibility analysis. The weightedMotifAccess package uses weight models to enable motif accessibility analysis, including transcription factor footprint information. The betterChromVAR package provides a novel, analytical re-implementation of the popular chromVAR method that offers substantial speed improvements, eliminates stochasticity, and offers additional features. Based on this, we also propose a method, CVnorm, that outperforms alternatives in normalizing technical bias in peak count data. The computational efficiency of these tools further enables a new framework for systematically investigating synergistic and antagonistic interactions between transcription factor motifs. Finally, the epiwraps package streamlines the visualization, normalization, and summarization of epigenomic data.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- simATAC: A single-cell ATAC-seq simulation framework 93%
- Estimating DNA-DNA interaction frequency from Hi-C data at restriction-fragment resolution 93%
- Simultaneous smoothing and detection of topological units of genome organization from sparse chromatin contact count matrices with matrix factorization 93%
Similar papers in this journal
- Semi-parametric Empirical Bayes Method for Multiplet Detection in snATAC-seq with Probabilistic Multi-omic Integration 93%
- magpie: a power evaluation method for differential RNA methylation analysis in N6-methyladenosine sequencing 93%
- Identifying promoter sequence architectures via a chunking-based algorithm using non-negative matrix factorisation 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.