LuxHS: DNA methylation analysis with spatially varying correlation structure
Halla-aho, V.; Lähdesmäki, H.
Show abstract
Bisulfite sequencing (BS-seq) is a popular method for measuring DNA methylation in basepair-resolution. Many BS-seq data analysis tools utilize the assumption of spatial correlation among the neighboring cytosines methylation states. While being a fair assumption, most existing methods leave out the possibility of deviation from the spatial correlation pattern. Our approach builds on a method which combines a generalized linear mixed model (GLMM) with a likelihood that is specific for BS-seq data and that incorporates a spatial correlation for methylation levels. We propose a novel technique using a sparsity promoting prior to enable cytosines deviating from the spatial correlation pattern. The method is tested with both simulated and real BS-seq data and compared to other differential methylation analysis tools.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- LuxUS: DNA methylation analysis using generalized linear mixed model with spatial correlation 97%
- interpolatedXY: a two-step strategy to normalise DNA methylation microarray data avoiding sex bias 95%
- MethPanel: a parallel pipeline and interactive analysis tool for multiplex bisulphite PCR sequencing to assess DNA methylation biomarker panels for disease detection 95%
Similar papers in this journal
- Predicting Differentially Methylated Cytosines in TET and DNMT3 Knockout Mutants via a Large Language Model 93%
- Molecular Group and Correlation Guided Structural Learning for Multi-Phenotype Prediction 93%
- Comparing full variation profile analysis with the conventional consensus method in SARS-CoV-2 phylogeny 93%
Similar papers in this journal
- DNA methylation-based sex classifier to predict sex and identify sex chromosome aneuploidy 94%
- Hierarchical non-negative matrix factorization using clinical information for microbial communities. 92%
- lncDIFF: a novel distribution-free method for differential expression analysis of long non-coding RNA 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.