Exploiting Electrophysiological Measures of Semantic Processing for Auditory Attention Decoding
Dijkstra, K.; Desain, P.; Farquhar, J.
Show abstract
In Auditory Attention Decoding, a users electrophysiological brain responses to certain features of speech are modelled and subsequently used to distinguish attended from unattended speech in multi-speaker contexts. Such approaches are frequently based on acoustic features of speech, such as the auditory envelope. A recent paper shows that the brains response to a semantic description (i.e., semantic dissimilarity) of narrative speech can also be modelled using such an approach. Here we use the (publicly available) data accompanying that study, in order to investigate whether combining this semantic dissimilarity feature with an auditory envelope approach improves decoding performance over using the envelope alone. We analyse data from their Cocktail Party experiment in which 33 subjects attended to one of two simultaneously presented audiobook narrations, for 30 1-minute fragments. We find that the addition of the dissimilarity feature to an envelope-based approach significantly increases accuracy, though the increase is marginal (85.4% to 86.6%). However, we subsequently show that this dissimilarity feature, in which the degree of dissimilarity of the current word with regard to the previous context is tagged to the onsets of each content word, can be replaced with a binary content-word-onset feature, without significantly affecting the results (i.e., modelled responses or accuracy), putting in question the added value of the dissimilarity information for the approach introduced in this recent paper.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Bayesian Prior Uncertainty and Surprisal Elicit Distinct Neural Patterns During Sound Localization in Dynamic Environments 96%
- Sound perception in realistic surgery scenarios: Towards EEG-based auditory work strain measures for medical personnel. 96%
- Audio-visual combination of syllables involves time-sensitive dynamics following from fusion failure 96%
Similar papers in this journal
- Neural representation of linguistic feature hierarchy reflects second-language proficiency 97%
- Post-hoc modification of linear models: combining machine learning with domain information to make solid inferences from noisy data 95%
- Decoding of selective attention to continuous speech from the human auditory brainstem response 94%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.