Back

CoAtGIN: Marrying Convolution and Attention for Graph-based Molecule Property Prediction

Cui, X.

2022-08-29 bioinformatics
10.1101/2022.08.26.505499 bioRxiv
Show abstract

Molecule property prediction based on computational strategy plays a key role in the process of drug discovery and design, such as DFT. Yet, these traditional methods are time-consuming and labour-intensive, which cant satisfy the need of biomedicine. Thanks to the development of deep learning, there are many variants of Graph Neural Networks (GNN) for molecule representation learning. However, whether the existed well-perform graph-based methods have a number of parameters, or the light models cant achieve good grades on various tasks. In order to manage the trade-off between efficiency and performance, we propose a novel model architecture, CoAtGIN, using both Convolution and Attention. On the local level, k-hop convolution is designed to capture long-range neighbour information. On the global level, besides using the virtual node to pass identical messages, we utilize linear attention to aggregate global graph representation according to the importance of each node and edge. In the recent OGB Large-Scale Benchmark, CoAtGIN achieves the 0.0933 Mean Absolute Error (MAE) on the large-scale dataset PCQM4Mv2 with only 5.6 M model parameters. Moreover, using the linear attention block improves the performance, which helps to capture the global representation.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
Briefings in Bioinformatics
354 papers in training set
Top 0.2%
18.4%
2
Bioinformatics
1204 papers in training set
Top 2%
12.6%
3
IEEE Transactions on Computational Biology and Bioinformatics
20 papers in training set
Top 0.1%
9.7%
4
IEEE/ACM Transactions on Computational Biology and Bioinformatics
38 papers in training set
Top 0.1%
6.7%
5
IEEE Journal of Biomedical and Health Informatics
37 papers in training set
Top 0.2%
4.8%
50% of probability mass above
6
PLOS Computational Biology
1863 papers in training set
Top 9%
4.0%
7
Journal of Chemical Information and Modeling
238 papers in training set
Top 1%
4.0%
8
BMC Bioinformatics
457 papers in training set
Top 3%
3.2%
9
Computational and Structural Biotechnology Journal
242 papers in training set
Top 2%
2.4%
10
PLOS ONE
5266 papers in training set
Top 45%
2.1%
11
Scientific Reports
3612 papers in training set
Top 50%
2.0%
12
Journal of Computational Biology
48 papers in training set
Top 0.5%
1.9%
13
Expert Systems with Applications
11 papers in training set
Top 0.1%
1.7%
14
Patterns
78 papers in training set
Top 1%
1.7%
15
IEEE Access
35 papers in training set
Top 0.9%
1.3%
16
iScience
1154 papers in training set
Top 26%
1.1%
17
Bioinformatics Advances
203 papers in training set
Top 4%
1.1%
18
Computers in Biology and Medicine
128 papers in training set
Top 3%
1.1%
19
Computational Biology and Chemistry
28 papers in training set
Top 0.8%
1.1%
20
Frontiers in Computational Neuroscience
60 papers in training set
Top 1%
1.0%
21
Nature Communications
5641 papers in training set
Top 57%
0.8%
22
Science Bulletin
21 papers in training set
Top 0.5%
0.6%
23
Genomics, Proteomics & Bioinformatics
172 papers in training set
Top 2%
0.6%
24
Journal of Cheminformatics
29 papers in training set
Top 0.8%
0.6%