An Explainable AI-Driven Classification Framework for Parkinson's Disease Detection via Acoustic Speech Features: A Comparative Machine Learning Study
Howlader, D.; AHMED, T.; Rahman, M. M.
Show abstract
Parkinsons Disease (PD) is a progressive neurodegenerative disorder which significantly affects motor function, daily coordination and verbal communication. Speech-based biomarkers provide a non-invasive and scalable approach to early detection, as dysphonia is one of the earliest and most consistent clinical markers of PD. The dataset used in this study is publicly available and consists of 756 voice recordings from 252 subjects (188 with PD and 64 neurologically healthy controls) with a wide range of acoustic parameters such as Mel-Frequency Cepstral Coefficients (MFCCs), energy-based parameters, and higher-order statistical derivatives. After systematic preprocessing and z-score normalisation, five machine learning classifiers were tested: K-Nearest Neighbors (KNN), Extreme Gradient Boosting (XGBoost), Random Forest (RF), Support Vector Machine (SVM), and Naive Bayes (NB) under a subject-independent, GroupKFold cross-validation protocol. The KNN classifier performed best overall with an accuracy of 92.10%, F1 score of 94.50%, and a precision rate of 98.09%, reducing the number of false positive diagnoses. To overcome the lack of interpretability of black-box predictive models, SHapley Additive exPlanations (SHAP) were used to explain the contribution of each feature to the prediction of an individual. The most diagnostically salient acoustic biomarkers were identified as features from the SHAP analysis: std delta delta log energy, the first Mel-Frequency Cepstral Coefficient, and Tunable Q-Factor Wavelet Transform (TQWT). This work introduces a machine learning framework that is both reproducible and clinically interpretable, combining high predictive accuracy with transparent, physiologically grounded decision logic.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Analyzing wav2vec embedding in Parkinson’s disease speech: A study on cross-database classification and regression tasks 96%
- Trends in Technology Usage for Parkinson's Disease Assessment: A Systematic Review 96%
- Development of a Tremor Detection Algorithm for use in an Academic Movement Disorders Center 95%
Similar papers in this journal
- Limited diagnostic accuracy of smartphone-based digital biomarkers for Parkinson’s disease in a remotely-administered setting 97%
- Sensitive quantification of cerebellar speech abnormalities using deep learning models 94%
- SleepSatelightFTC: A Lightweight and Interpretable Deep Learning Model for Single-Channel EEG-Based Sleep Stage Classification 92%
Similar papers in this journal
- Autocorrelation-based method to identify disordered rhythm in Parkinsons disease tasks: a novel approach applicable to multimodal devices 95%
- Quantifying normal and parkinsonian gait features from home movies: Practical application of a deep learning-based 2D pose estimator 94%
- Comparing P300 flashing paradigms in online typing with language models 93%
Similar papers in this journal
- Multimodal Speech Biomarkers for Remote Monitoring of ALS Disease Progression 95%
- SymScore: Machine Learning Accuracy Meets Transparency in a Symbolic Regression-Based Clinical Score Generator 92%
- A machine-learning Approach for Stress Detection Using Wearable Sensors in Free-living Environments 92%
Similar papers in this journal
- Identification of predictive patient characteristics for assessing the probability of COVID-19 in-hospital mortality 93%
- A recurrent neural network and parallel hidden Markov model algorithm to segment and detect heart murmurs in phonocardiograms 92%
- From theoretical models to practical deployment: A perspective and case study of opportunities and challenges in AI-driven healthcare research for low-income settings 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.