Genome Wide Analysis of Dinucleotide Distribution along the Genomes of Species and its Biological Implication
Cai, Z.; Liu, S.; Xue, Y.; Quan, H.; Zhang, L.; Gao, Y. Q.
Show abstract
Dinucleotide densities and their distribution patterns vary significantly among species. Previous studies revealed that CpG is susceptible to methylation, enriched at topologically associating domains (TADs) boundaries and its distribution along the genome correlates with chromatin compartmentalization. However, the multi-scale organizations of CpG in the linear genome, their role in chromatin organization, and how they change along the evolution are only partially understood. By comparing the CpG distribution at different genomic length scales, we quantify the difference between the CpG distributions of different species and evaluate how the hierarchical uneven CpG distribution appears in evolution. The clustering of species based on the CpG distribution is consistent with the phylogenetic tree. Interestingly, we found the CpG distribution and chromatin structure to be correlated in many different length scales, especially for mammals and avians, consistent with the mosaic CpG distribution in the genomes of these species.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Reconstruction of segmental duplication rates and associated genomic features by network analysis 93%
- Evaluating chromatin accessibility differences across multiple primate species using a joint modelling approach 93%
- Comparative genomics analysis reveals high levels of differential DNA transposition among primates 92%
Similar papers in this journal
- A map of cis-regulatory modules and constituent transcription factor binding sites in 80% of the mouse genome 95%
- Microsatellite Density Landscapes Illustrate Short Tandem Repeats Aggregation in The Complete Reference Human Genome 93%
- Deciphering epigenomic code for cell differentiation using deep learning 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.