Imagined speech can be decoded from low- and cross-frequency features in perceptual space
Proix, T.; Delgado Saa, J.; Christen, A.; Martin, S.; Pasley, B. N.; Knight, R. T.; Tian, X.; Poeppel, D.; Doyle, W. K.; Devinsky, O.; Arnal, L. H.; Megevand, P.; Giraud, A.-L.
Show abstract
Reconstructing intended speech from neural activity using brain-computer interfaces (BCIs) holds great promises for people with severe speech production deficits. While decoding overt speech has progressed, decoding imagined speech have met limited success, mainly because the associated neural signals are weak and variable hence difficult to decode by learning algorithms. Using three electrocorticography datasets totalizing 1444 electrodes from 13 patients who performed overt and imagined speech production tasks, and based on recent theories of speech neural processing, we extracted consistent and specific neural features usable for future BCIs, and assessed their performance to discriminate speech items in articulatory, phonetic, vocalic, and semantic representation spaces. While high-frequency activity provided the best signal for overt speech, both low- and higher-frequency power and local cross-frequency contributed to successful imagined speech decoding, in particular in phonetic and vocalic, i.e. perceptual, spaces. These findings demonstrate that low-frequency power and cross-frequency dynamics contain key information for imagined speech decoding, and that exploring perceptual spaces offers a promising avenue for future imagined speech BCIs.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Processing of auditory feedback in perisylvian and insular cortex 98%
- Differential auditory and visual phase-locking are observed during audio-visual benefit and silent lip-reading for speech perception 97%
- Dimensionality and ramping: Signatures of sentence integration in the dynamics of brains and deep language models 97%
Similar papers in this journal
- Attentional Modulation of Hierarchical Speech Representations in a Multitalker Environment 97%
- Expectations boost the reconstruction of auditory features from electrophysiological responses to noisy speech 96%
- Two Stages of Speech Envelope Tracking in Human Auditory Cortex Modulated by Speech Intelligibility 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.