Back

Estimating the Explainable Variance of EEG Responses to Natural Speech

Dou, J.; Lalor, E.

2026-07-03 neuroscience
10.64898/2026.07.02.736170 bioRxiv
Show abstract

Substantial progress has been made in recent years on understanding how the human brain parses and processes natural speech. Much of this progress has been based on modeling how brain activity relates to the different acoustic and linguistic features of speech. By fitting and testing models based on those features, one can test hypotheses about the kinds of computations and representations the brain uses to convert speech sounds into understanding. While much of this work has focused on modeling BOLD activity using functional neuroimaging or intracranially recorded electrophysiological signals, the approach has also proven useful with MEG and EEG. Indeed, noninvasive EEG has certain advantages for studying speech processing in terms of translational research and application. Research over the last decade or so has shown that EEG can be successfully modeled based on numerous acoustic, linguistic, and paralinguistic speech features. However, an important unanswered question hangs over all of this work: namely, what constitutes a good model of EEG responses to natural speech? Or, to put it another way, how much variance in EEG recorded during natural speech listening is explainable as having derived from that speech input? The present study aims to tackle this issue. We do so under the assumption that the best model for a person's EEG response to natural speech is a set of EEG responses from other people listening to the same speech. Using this assumption, we construct inter-subject models using EEG from 19 healthy adult native speakers of English who all listened to the same audiobook. The model for each subject involves predicting their EEG data using (dimensionality-reduced) EEG from different numbers of other subjects and then extrapolating to estimate the total explainable variance in the target individual's response to speech. Following this, we show that linear models (temporal response functions) based on several commonly used acoustic and linguistic speech features can predict most - but importantly not all - of the estimated total explainable variance in EEG responses across subjects.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
NeuroImage
903 papers in training set
Top 0.3%
26.7%
2
Frontiers in Human Neuroscience
77 papers in training set
Top 0.1%
7.9%
3
Scientific Reports
3612 papers in training set
Top 13%
6.3%
4
European Journal of Neuroscience
189 papers in training set
Top 0.4%
5.5%
5
PLOS Computational Biology
1863 papers in training set
Top 6%
5.5%
50% of probability mass above
6
Frontiers in Neuroscience
256 papers in training set
Top 0.4%
5.5%
7
PLOS ONE
5266 papers in training set
Top 31%
4.9%
8
Brain Topography
29 papers in training set
Top 0.1%
3.2%
9
Journal of Neural Engineering
221 papers in training set
Top 1%
2.4%
10
Journal of Neuroscience Methods
122 papers in training set
Top 0.8%
2.4%
11
Journal of Neurophysiology
302 papers in training set
Top 2%
2.1%
12
Neural Computation
39 papers in training set
Top 0.4%
1.9%
13
eneuro
439 papers in training set
Top 5%
1.7%
14
Cerebral Cortex
396 papers in training set
Top 4%
1.3%
15
Hearing Research
54 papers in training set
Top 0.3%
1.3%
16
The Journal of Neuroscience
1025 papers in training set
Top 8%
1.1%
17
Imaging Neuroscience
282 papers in training set
Top 3%
1.1%
18
Journal of Cognitive Neuroscience
135 papers in training set
Top 2%
1.0%
19
Proceedings of the National Academy of Sciences
2444 papers in training set
Top 41%
0.8%
20
eLife
5828 papers in training set
Top 64%
0.8%
21
Psychophysiology
77 papers in training set
Top 1%
0.8%
22
Human Brain Mapping
329 papers in training set
Top 4%
0.8%
23
Biomedical Signal Processing and Control
22 papers in training set
Top 0.7%
0.8%
24
Cognitive Neurodynamics
18 papers in training set
Top 0.5%
0.6%
25
Nature Communications
5641 papers in training set
Top 59%
0.6%