The SEQC2 Epigenomics Quality Control (EpiQC) Study: Comprehensive Characterization of Epigenetic Methods, Reproducibility, and Quantification
Foox, J.; Nordlund, J.; Lalancette, C.; Gong, T.; Lacey, M.; Lent, S.; Langhorst, B. W.; Ponnaluri, V. K. C.; Williams, L.; Padmamabhan, K.; Cavalcante, R. G.; Lundmark, A.; Butler, D.; Gurvitch, J. M.; Greally, J. M.; Suzuki, M.; Menor, M.; Nasu, M.; Alonso, A.; Sheridan, C.; Scherer, A.; Bruinsma, S.; Golda, G.; Muszynska, A.; Łabaj, P. P.; Campbell, M. A.; Wos, F.; Raine, A.; Liljedahl, U.; Axelsson, T.; Wang, C.; Chen, Z.; Yang, Z.; Li, J.; Yang, X.; Wang, H.; Melnick, A.; Guo, S.; Blume, A.; Franke, V.; Ibanez de Caceres, I.; Rodriguez-Antolin, C.; Rosas, R.; Davis, J. W.; Ishii, J.; Megh
Show abstract
Cytosine modifications in DNA such as 5-methylcytosine (5mC) underlie a broad range of developmental processes, maintain cellular lineage specification, and can define or stratify cancer and other diseases. However, the wide variety of approaches available to interrogate these modifications has created a need for harmonized materials, methods, and rigorous benchmarking to improve genome-wide methylome sequencing applications in clinical and basic research. Here, we present a multi-platform assessment and a global resource for epigenetics research from the FDAs Epigenomics Quality Control (EpiQC) Group. The study design leverages seven human cell lines that are designated as reference materials and publicly available from the National Institute of Standards and Technology (NIST) and Genome in a Bottle (GIAB) consortium. These samples were subject to a variety of genome-wide methylation interrogation approaches across six independent laboratories, with a primary focus was on 5-methylcytosine modifications. Each sample was processed in two or more technical replicates by three whole-genome bisulfite sequencing (WGBS) protocols (TruSeq DNA methylation, Accel-NGS MethylSeq, and SPLAT), oxidative bisulfite sequencing (TrueMethyl), one enzymatic deamination method (EMseq), targeted methylation sequencing (Illumina Methyl Capture EPIC), and single-molecule long-read nanopore sequencing from Oxford Nanopore Technologies. After rigorous quality assessment and comparison to Illumina EPIC methylation microarrays and testing on a range of algorithms (Bismark, BitmapperBS, BWAMeth, and GemBS), we found overall high concordance between assays (R=0.87-R0.93), differences in efficency of read mapping and CpG capture and coverage, and platform performance. The data provided herein can guide continued used of these reference materials in epigenomics assays, as well as provide best practices for epigenomics research and experimental design in future studies.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Non-destructive enzymatic deamination enables single molecule long read sequencing for the determination of 5-methylcytosine and 5-hydroxymethylcytosine at single base resolution. 97%
- EM-seq: Detection of DNA Methylation at Single Base Resolution from Picograms of DNA 96%
- Simultaneous assessment of human genome and methylome data in a single experiment using limited deamination of methylated cytosine. 95%
Similar papers in this journal
- A mammalian methylation array for profiling methylation levels at conserved sequences 97%
- Systematic benchmarking of tools for CpG methylation detection from Nanopore sequencing 96%
- MethylBERT: A Transformer-based model for read-level DNA methylation pattern identification and tumour deconvolution 94%
Similar papers in this journal
- Permutation-based significance analysis reduces the type 1 error rate in bisulfite sequencing data analysis of human umbilical cord blood samples 94%
- A Cell Type Enrichment Analysis Tool for Brain DNA Methylation Data (CEAM) 94%
- Whole human genome 5'-mC methylation analysis using long read nanopore sequencing 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.