Automated Detection of Speech Disorders in Parkinson's Disease using Deep Convolutional Neural Networks: A Pilot Study
Jones, S.; Cosgrove, J.; Wang, H.; Mathew, R.
Show abstract
BackgroundPatients with Parkinsons disease (PD) frequently exhibit deficits in functional communication due to the presence of speech disorders associated with dysarthria that can be characterized by monotony of pitch (or fundamental frequency), reduced loudness, irregular rate of speech, imprecise consonants, and changes in voice quality. This pilot study investigates the application of a speech classifier based on deep-convolutional neural networks (CNNs) for aiding early diagnosis of PD. MethodsIn this study, we analyse the performance capabilities of two audio feature extraction techniques and associated model architectures: low-level time-frequency based features classified using a Support Vector Machine (SVM); and classifying log mel-spectrograms of segmented audio signals using varying depths of Deep Convolutional Neural Networks. The models were trained using an open-source data set comprised of 73 audio recordings of continuous dialogue from 37 subjects, including 16 people with PD (5 females and 11 males) and 21 healthy controls (19 females and 2 males), who were required to perform two speech production tasks. ResultsThe experimental results show that the deep CNN model, trained on the log mel-spectrograms of 5-second segmented audio signals, can successfully differentiate PD subjects from healthy controls (HC) with a mean accuracy of 84.7%, sensitivity of 87.9% sensitivity and specificity of 89.4%, thus demonstrating its potential for aiding early diagnosis of PD in a clinical setting. The saliency maps show that the deep CNN model can distinguish between PD participants and healthy controls by detecting centralised, low-frequency regions of the spectrograms representing the speech of PD subjects, whereas a larger range of frequencies are detected in the spectrograms representing speech from healthy controls.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Analyzing wav2vec embedding in Parkinson’s disease speech: A study on cross-database classification and regression tasks 98%
- Development of a Tremor Detection Algorithm for use in an Academic Movement Disorders Center 94%
- Motor signatures in digitized cognitive and memory tests enhances characterization of Parkinson’s disease 93%
Similar papers in this journal
- Sensitive quantification of cerebellar speech abnormalities using deep learning models 96%
- Limited diagnostic accuracy of smartphone-based digital biomarkers for Parkinson’s disease in a remotely-administered setting 94%
- SleepSatelightFTC: A Lightweight and Interpretable Deep Learning Model for Single-Channel EEG-Based Sleep Stage Classification 91%
Similar papers in this journal
- Autocorrelation-based method to identify disordered rhythm in Parkinsons disease tasks: a novel approach applicable to multimodal devices 94%
- Comparing P300 flashing paradigms in online typing with language models 93%
- A Signal Demodulation-based Method for the Early Detection of Cheyne-Stokes Respiration 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.