Distinct neurolinguistic signatures of patients with acquired neurological conditions:. An open library of clinically relevant linguistic biomarkers
Themistocleous, C.; Stark, B. C.
Show abstract
Individuals with left-hemisphere damage (LHD), right-hemisphere damage (RHD), dementia, mild cognitive impairment (MCI), traumatic brain injury (TBI), and healthy controls are characterized by overlapping clinical profiles affecting communication and social interaction. Language provides a rich, non-invasive window into neurological health, yet objective and scalable methods to automatically differentiate between conditions with are lacking. This method aims to develop comprehensive neurolinguistic measures of these conditions, develop a machine learning multiclass screening and language assessment model (NeuroScreen) and offer a large comparative database of measures for other studies to build upon. We combined one of the largest databases, comprising 291 linguistic biomarkers calculated from speech samples produced by 1,394 participants: 536 individuals with aphasia secondary to LHD, 193 individuals with dementia, 107 individuals with MCI, 38 individuals with RHD, 58 individuals with TBI, and 498 Healthy Controls. Employing natural language processing (NLP) via the Open Brain AI platform (http://openbrainai.com), we extracted multiple linguistic features from the speech samples, including readability, lexical richness, phonology, morphology, syntax, and semantics. A Deep Neural Network architecture (DNN) classifies these conditions from linguistic features with high accuracy (up to 91%). A linear mixed-effects model approach was employed to determine the biomarkers of the neurological conditions, revealing distinct, quantitative neurolinguistic properties: LHD and TBI show widespread deficits in syntax and phonology; MCI is characterized by fine-grained simplification; patients with dementia present with specific lexico-semantic impairments; and RHD shows the most preserved profile. Ultimately, the outcomes provide an automatic detection and classification model of key neurological conditions affecting language, along with a novel set of validated neurological markers for facilitating differential diagnosis, remote monitoring, and personalized neurological care.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Lexical Markers of Disordered Speech in Primary Progressive Aphasia and ‘Parkinson-plus’ Disorders 96%
- Verbal fluency tests assess global cognitive status but have limited diagnostic differentiation: Evidence from a large-scale examination of six neurodegenerative diseases 95%
- Impaired semantic control in the logopenic variant of primary progressive aphasia 94%
Similar papers in this journal
- Cognitive neuropsychological and neuroanatomic predictors of naturalistic action performance in left hemisphere stroke: a retrospective analysis 95%
- Establishing and evaluating the gradient of item naming difficulty in post-stroke aphasia and semantic dementia 95%
- Fluid intelligence and naturalistic task impairments after focal brain lesions 94%
Similar papers in this journal
Similar papers in this journal
- Going off the rails: Impaired coherence in the speech of patients with semantic control deficits 94%
- Interacting effects of frontal lobe neuroanatomy and working memory capacity to older listeners' speech recognition in noise 93%
- Perceptual and semantic deficits in face recognition in semantic dementia 93%
Similar papers in this journal
- Kinematic Correlates of Early Speech Motor Changes in Cognitively Intact APOE-ε4 Carriers: A Preliminary Study Using a Color-Word Interference Task 92%
- Impaired everyday executive functions and cognitive strategy use on the Weekly Calendar Planning Activity in individuals with stroke undergoing acute inpatient rehabilitation 91%
- Ontario Neurodegenerative Disease Research Initiative (ONDRI): Structural MRI methods & outcome measures 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.