Back

Large-scale integration of DNA methylation and gene expression array platforms

Lancaster, E. E.; Vladimirov, V. I.; Riley, B. P.; Landry, J. W.; Roberson-Nay, R.; York, T. P.

2021-08-12 genetics
10.1101/2021.08.12.455267 bioRxiv
Show abstract

AO_SCPLOWBSTRACTC_SCPLOWEpigenome-wide association studies (EWAS) aim to provide evidence that marks of DNA methylation (DNAm) have downstream consequences that can result in the development of human diseases. Although these methods have been successful in identifying DNAm patterns associated with disease states, any further characterization of etiologic mechanisms underlying disease remains elusive. This knowledge gap does not originate from a lack of DNAm-trait associations, but rather stems from study design issues that affect the interpretability of EWAS results. Despite known limitations in predicting the function of a particular CpG site, most EWAS maintain the broad assumption that altered DNAm results in a concomitant change of transcription at the most proximal gene. This study integrated DNAm and gene expression (GE) measurements in two cohorts, the Adolescent and Young Adult Twin Study (AYATS) and the Pregnancy, Race, Environment, Genes (PREG) study, to improve the understanding of epigenomic regulatory mechanisms. CpG sites associated with GE in cis were enriched in areas of transcription factor binding and areas of intermediate-to-low CpG density. CpG sites associated with trans GE were also enriched in areas of known regulatory significance, including enhancer regions. These results highlight issues with restricting DNAm-transcript annotations to small genomic intervals and question the validity of assuming a canonical cis DNAm-GE pathway. Based on these findings, the interpretation of EWAS results is limited in studies without multi-omic support and further research should identify genomic regions in which GE-associated DNAm is overrepresented.

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.