Detecting Primary Progressive Aphasia (PPA) from Text: A Benchmarking Study
Merhbene, G.; Lecron, F.; Fortemps, P.; Dickerson, B. C.; Kurpicz-Briki, M.; Rezaii, N.
Show abstract
Classifying subtypes of primary progressive aphasia (PPA) from connected speech presents significant diagnostic challenges due to overlapping linguistic markers. This study benchmarks the performance of traditional machine learning models with various feature extraction techniques, transformer-based models, and large language models (LLMs) for PPA classification. Our results indicate that while transformerbased models and LLMs exceed chance-level performance in terms of balanced accuracy, traditional classifiers combined with contextual embeddings remain highly competitive. Notably, SVM using RoBERTas embeddings achieves the highest classification accuracy. These findings underscore the potential of machine learning in enhancing the automatic classification of PPA subtypes.
Matching journals
The top 11 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Building Large-Scale Registries from Unstructured Clinical Notes using a Low-Resource Natural Language Processing Pipeline 92%
- Uncertainty in Deep Learning for EEG under Dataset Shifts 92%
- Deep ensemble multitask classification of emergency medical call incidents combining multimodal data improves emergency medical dispatch 91%
Similar papers in this journal
- Tracking lexical and semantic prediction error underlying the N400 using artificial neural network models of sentence processing 91%
- Neural correlates of object-extracted relative clause processing across English and Chinese 89%
- Assessing the sensitivity of EEG-based frequency-tagging as a metric for statistical learning 89%
Similar papers in this journal
- Regularized Bagged Canonical Component Analysis for Multiclass Learning in Brain Imaging 92%
- Towards Multi-Brain Decoding in Autism: A Self-Supervised Learning Approach 90%
- Building Models of Functional Interactions Among Brain Domains that Encode Varying Information Complexity: A Schizophrenia Case Study 90%
Similar papers in this journal
- Active Neural Networks to Detect Mentions of Changes to Medication Treatment in Social Media 94%
- LCD Benchmark: Long Clinical Document Benchmark on Mortality Prediction for Language Models 93%
- Annotation-preserving machine translation of English corpora to validate Dutch clinical concept extraction tools 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.