Interpretable multi-timescale models for predicting fMRI responses to continuous natural speech
Jain, S.; Vo, V. A.; Mahto, S.; LeBel, A.; Turek, J. S.; Huth, A. G.
Show abstract
Natural language contains information at multiple timescales. To understand how the human brain represents this information, one approach is to build encoding models that predict fMRI responses to natural language using representations extracted from neural network language models (LMs). However, these LM-derived representations do not explicitly separate information at different timescales, making it difficult to interpret the encoding models. In this work we construct interpretable multi-timescale representations by forcing individual units in an LSTM LM to integrate information over specific temporal scales. This allows us to explicitly and directly map the timescale of information encoded by each individual fMRI voxel. Further, the standard fMRI encoding procedure does not account for varying temporal properties in the encoding features. We modify the procedure so that it can capture both short- and long-timescale information. This approach outperforms other encoding models, particularly for voxels that represent long-timescale information. It also provides a finer-grained map of timescale information in the human language pathway. This serves as a framework for future work investigating temporal hierarchies across artificial and biological language systems.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Dynamic Network Analysis of Electrophysiological Task Data 95%
- Alignment massive of auditory individual artificial networks with fMRI brain data leads to generalizable improvements in brain encoding and downstream tasks 95%
- Spectral graph model for fMRI: a biophysical, connectivity-based generative model for the analysis of frequency-resolved resting state fMRI 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.