Back

Creation and Validation of the First Infinium DNA Methylation Array for the Human Imprintome

Carreras-Gallo, N.; Dwaraka, V. B.; Jima, D. D.; Skaar, D. A.; Mendez, T. L.; Planchart, A.; Zhou, W.; Jirtle, R. L.; Smith, R.; Hoyo, C.

2024-01-16 genetics
10.1101/2024.01.15.575646 bioRxiv
Show abstract

BackgroundDifferentially methylated imprint control regions (ICRs) regulate the monoallelic expression of imprinted genes. Their epigenetic dysregulation by environmental exposures throughout life results in the formation of common chronic diseases. Unfortunately, existing Infinium methylation arrays lack the ability to profile these regions adequately. Whole genome bisulfite sequencing (WGBS) is the unique method able to profile these regions, but it is very expensive and it requires not only a high coverage but it is also computationally intensive to assess those regions. FindingsTo address this deficiency, we developed a custom methylation array containing 22,819 probes. Among them, 9,757 probes map to 1,088 out of the 1,488 candidate ICRs recently described. To assess the performance of the array, we created matched samples processed with the Human Imprintome array and WGBS, which is the current standard method for assessing the methylation of the Human Imprintome. We compared the methylation levels from the shared CpG sites and obtained a mean R2 = 0.569. We also created matched samples processed with the Human Imprintome array and the Infinium Methylation EPIC v2 array and obtained a mean R2 = 0.796. Furthermore, replication experiments demonstrated high reliability (ICC: 0.799-0.945). ConclusionsOur custom array will be useful for replicable and accurate assessment, mechanistic insight, and targeted investigation of ICRs. This tool should accelerate the discovery of ICRs associated with a wide range of diseases and exposures, and advance our understanding of genomic imprinting and its relevance in development and disease formation throughout the life course.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.