An Innovative Framework for Heart Sound Classification Integrating Adaptive Fuzzy Rank-Based Ensemble of Transfer Learning Models with Multi-Dimensional Features
Sarker, S.; Fairuj Preotee, F.; Muhammad, T.; Akhter, S.
Show abstract
Heart sound identification for diagnosing cardiovascular diseases is a complex challenge due to the intricate, spectro-temporal characteristics of conditions like Aortic Stenosis, Mitral Regurgitation, and multi-valvular disorders. Conventional methods frequently inadequately encompass the complete range of diagnostic characteristics, depending on fragmented or simplistic approaches that lack generalizability across varied patient demographics, clinical settings, and recording circumstances. To address these limitations, we offer an innovative, multi-dimensional feature fusion framework that integrates Wavelet Scattering Transform (WST) and Mel-Frequency Cepstral Coefficients (MFCC) to capture both temporal stability and perceptually optimal frequency patterns from cardiac sounds. This method is further refined by an adaptive fuzzy rank-based ensemble technique utilizing the Gompertz function, which dynamically modifies model weights according to confidence metrics, hence assuring more dependable and precise predictions amidst fluctuating clinical uncertainty. We meticulously assess our model utilizing eight advanced fine-tuned transfer learning architectures across four feature extraction techniques (WST, MFCC, STFT, Multi-Dimensional) on the clinically validated BUET Multi-disease Heart Sound (BMD-HS) dataset. This dataset comprises 864 phonocardiogram recordings from 108 participants with echocardiographically confirmed diagnoses. The multi-dimensional feature fusion method attains 97% accuracy with EfficientNetB2, whereas the adaptive fuzzy ensemble strategy achieves 98% accuracy, surpassing both individual models and conventional ensemble methods. Moreover, Explainable AI with Audio-LIME offers transparent, clinically interpretable insights, achieving fidelity scores surpassing 0.85 and clinical relevance ratings exceeding 90%, facilitating the identification of critical time-frequency regions that hold diagnostic significance. This system establishes a novel benchmark for heart sound classification, enhances the differentiation of intricate valvular disorders, and provides dependable, evidence-based decision support for cardiovascular diagnosis in clinical settings. Our research illustrates the capability for resilient, scalable, and interpretable heart sound classification systems that can improve clinical decision-making and promote the integration of automated diagnostic tools in healthcare.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Detecting Heart Failure using novel bio-signals and a Knowledge Enhanced Neural Network 95%
- A machine-learning Approach for Stress Detection Using Wearable Sensors in Free-living Environments 93%
- BenchXAI: Comprehensive Benchmarking of Post-hoc Explainable AI Methods on Multi-Modal Biomedical Data 93%
Similar papers in this journal
- Hilbert-Envelope Features for Cardiac Disease Classification from Noisy Phonocardiograms 97%
- Digital Voice-Based Biomarker for Monitoring Respiratory Quality of Life: Findings from the Colive Voice Study 95%
- Improved online event detection and differentiation by a simple gradient-based nonlinear transformation: Implications for the biomedical signal and image analysis 95%
Similar papers in this journal
Similar papers in this journal
- Rett syndrome severity estimation with the BioStamp nPoint using interactions between heart rate variability and body movement 95%
- A Signal Demodulation-based Method for the Early Detection of Cheyne-Stokes Respiration 94%
- 3 Directional Inception-ResUNet: deep spatial feature learning for multichannel singing voice separation with distortion 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.