Back

Reproducible comparison and interpretation of machine learning classifiers to predict autism on the ABIDE multimodal dataset

Dong, Y.; Batalle, D.; Deprez, M.

2024-09-04 neurology
10.1101/2024.09.04.24313055 medRxiv
Show abstract

Autism is a neurodevelopmental condition affecting [~]1% of the population. Recently, machine learning models have been trained to classify participants with autism using their neuroimaging features, though the performance of these models varies in the literature. Differences in experimental setup hamper the direct comparison of different machine-learning approaches. In this paper, five of the most widely used and best-performing machine learning models in the field were trained to classify participants with autism and typically developing (TD) participants, using functional connectivity matrices, structural volumetric measures and phenotypic information from the Autism Brain Imaging Data Exchange (ABIDE) dataset. Their performance was compared under the same evaluation standard. The models implemented included: graph convolutional networks (GCN), edge-variational graph convolutional networks (EV-GCN), fully connected networks (FCN), auto-encoder followed by a fully connected network (AE-FCN) and support vector machine (SVM). Our results show that all models performed similarly, achieving a classification accuracy around 70%. Our results suggest that different inclusion criteria, data modalities and evaluation pipelines rather than different machine learning models may explain variations in accuracy in published literature. The highest accuracy in our framework was obtained by an ensemble of GCN models trained on combination of functional MRI and structural MRI features, reaching classification accuracy of 72.2% and AUC = 0.78 on the test set. The combined structural and functional modalities exhibited higher predictive ability compared to using single modality features alone. Ensemble methods were found to be helpful to improve the performance of the models. Furthermore, we also investigated the stability of features identified by the different machine learning models using the SmoothGrad interpretation method. The FCN model demonstrated the highest stability selecting relevant features contributing to model decision making. Code available at: https://github.com/YilanDong19/Machine-learning-with-ABIDE.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
Neuroinformatics
46 papers in training set
Top 0.1%
22.7%
2
NeuroImage
903 papers in training set
Top 2%
6.8%
3
Frontiers in Neuroinformatics
41 papers in training set
Top 0.1%
6.3%
4
Brain Connectivity
25 papers in training set
Top 0.1%
6.3%
5
Human Brain Mapping
329 papers in training set
Top 1%
5.6%
6
PLOS Digital Health
106 papers in training set
Top 1.0%
5.6%
50% of probability mass above
7
Computers in Biology and Medicine
128 papers in training set
Top 0.8%
4.4%
8
Network Neuroscience
126 papers in training set
Top 0.5%
3.5%
9
NeuroImage: Clinical
144 papers in training set
Top 0.8%
3.5%
10
Frontiers in Neuroscience
256 papers in training set
Top 1%
3.3%
11
Brain Sciences
55 papers in training set
Top 0.2%
3.2%
12
Scientific Reports
3612 papers in training set
Top 39%
2.7%
13
Cerebral Cortex
396 papers in training set
Top 3%
2.1%
14
Developmental Cognitive Neuroscience
96 papers in training set
Top 0.7%
1.7%
15
PLOS ONE
5266 papers in training set
Top 49%
1.7%
16
IEEE Journal of Biomedical and Health Informatics
37 papers in training set
Top 0.8%
1.4%
17
Imaging Neuroscience
282 papers in training set
Top 3%
1.1%
18
IEEE Access
35 papers in training set
Top 0.9%
1.1%
19
Brain Structure and Function
93 papers in training set
Top 1%
0.9%
20
Sensors
43 papers in training set
Top 1%
0.9%
21
Journal of Medical Imaging
11 papers in training set
Top 0.4%
0.9%
22
Frontiers in Artificial Intelligence
20 papers in training set
Top 0.9%
0.6%
23
Brain Informatics
10 papers in training set
Top 0.3%
0.6%
24
GigaScience
212 papers in training set
Top 5%
0.6%