Appropriate data segmentation improves speech encoding models
Bialas, O.; Lalor, E. C.
Show abstract
In recent decades, research on the neural processing of speech and language increasingly investigated ongoing responses to continuously presented naturalistic speech, allowing researchers to ask interesting questions about different representations of speech and their relationships. This requires statistical models that can dissect different sources of variance occurring in the processing of naturalistic speech. One commonly used family of models are temporal response functions (TRFs) which can predict neural responses to speech as a weighted combination of different features and points in time. TRFs model the brain as a linear time-invariant (LTI) system whose responses can be characterized by constant transfer functions. This implicitly assumes that the underlying signals are stationary, varying to a fixed degree around a constant mean. However, continuous neural recordings commonly violate this assumption. Here, we use simulations and EEG recordings to investigate how non-stationarities affect TRF models for continuous speech processing. Our results suggest that non-stationarities may impair the performance of TRF models, but that this can be partially remedied by dividing the data into shorter segments that approximate stationarity.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Post-hoc modification of linear models: combining machine learning with domain information to make solid inferences from noisy data 96%
- ZapLine: a simple and effective method to remove power line artifacts 96%
- The relationship between frequency content and representational dynamics in the decoding of neurophysiological data 96%
Similar papers in this journal
- The Neural Response at the Fundamental Frequency of Speech is Modulated by Word-level Acoustic and Linguistic Information 96%
- Decoding continuous variables from event-related potential (ERP) data with linear support vector regression (SVR) using the Decision Decoding Toolbox (DDTBOX) 96%
- Source Localization Using Recursively Applied and Projected MUSIC with Flexible Extent Estimation 94%
Similar papers in this journal
- Disentangling signal and noise in neural responses through generative modeling 95%
- Alpha blocking and 1/fβ spectral scaling in resting EEG can be accounted for by a sum of damped alpha band oscillatory processes 95%
- Time-resolved dynamic computational modeling of human EEG recordings reveals gradients of generative mechanisms for the MMN response 95%
Similar papers in this journal
- Different methods to estimate the phase of neural rhythms agree, but only during times of low uncertainty 95%
- Prefrontal High Gamma in ECoG tags periodicity of musical rhythms in perception and imagination 94%
- Practical Bayesian Inference in Neuroscience: Or How I Learned To Stop Worrying and Embrace the Distribution 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.