Improving Cell-type-specific 3D Genome Architectures Prediction Leveraging Graph Neural Networks
Wang, R.; Ma, W.; Mohammadi, A. S.; Shahsavari, S.; Vosoughi, S.; Wang, X.
Show abstract
The mammalian genome organizes into complex three-dimensional structures, where interactions among chromatin regulatory elements play a pivotal role in mediating biological functions, highlighting the significance of genomic region interactions in biological research. Traditional biological sequencing techniques like HiC and MicroC, commonly employed to estimate these interactions, are resource-intensive and time-consuming, especially given the vast array of cell lines and tissues involved. With the advent of advanced machine learning (ML) methodologies, there has been a push towards developing ML models to predict genomic interactions. However, while these models excel in predicting interactions for cell lines similar to their training data, they often fail to generalize across distantly related cell lines or accurately predict interactions specific to certain cell lines. Identifying the potential oversight of excluding example genomic region interaction information from model inputs as a fundamental limitation, this paper introduces GRACHIP, a model rooted in graph neural network technology aiming to address this issue by incorporating detailed interaction information as a hint. Through extensive testing across various cell lines, GRACHIP not only demonstrates exceptional accuracy in predicting chromatin interaction intensity but showcases remarkable generalizability to cell lines not encountered during training. Consequently, GRACHIP emerges as a potent research tool, offering a viable alternative to conventional sequencing methods for analyzing the interactions and three-dimensional organization of mammalian genomes, thus alleviating the dependency on expensive and time-consuming biological sequencing techniques. It also offers an alternative way for researchers to investigate 3D chromatin interactions and simulate their changes in model systems to test their hypotheses.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Message Passing Framework for Precise Cell State Identification with scClassify2 96%
- Simultaneous smoothing and detection of topological units of genome organization from sparse chromatin contact count matrices with matrix factorization 96%
- Learning latent embedding of multi-modal single cell data and cross-modality relationship simultaneously 95%
Similar papers in this journal
- DNALONGBENCH: A Benchmark Suite for Long-Range DNA Prediction Tasks 97%
- CellFM: a large-scale foundation model pre-trained on transcriptomics of 100 million human cells 97%
- Hi-C-LSTM: Learning representations of chromatin contacts using a recurrent neural network identifies genomic drivers of conformation 97%
Similar papers in this journal
- DeepPHiC: Predicting promoter-centered chromatin interactions using a novel deep learning approach 97%
- NetTIME: a multitask and base-pair resolution framework for improved transcription factor binding site prediction 97%
- Graph Convolutional Networks for Epigenetic State Prediction Using Both Sequence and 3D Genome Data 96%
Similar papers in this journal
- Deep generative modeling and clustering of single cell Hi-C data 97%
- Graph Contrastive Learning of Subcellular-resolution Spatial Transcriptomics Improves Cell Type Annotation and Reveals Critical Molecular Pathways 96%
- DeepDRIM: a deep neural network to reconstruct cell-type-specific gene regulatory network using single-cell RNA-seq data 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.