Decoding Speech and Music Stimuli from the Frequency Following Response
Losorelli, S.; Kaneshiro, B.; Musacchia, G. A.; Blevins, N. H.; Fitzgerald, M. B.
Show abstract
The ability to differentiate complex sounds is essential for communication. Here, we propose using a machine-learning approach, called classification, to objectively evaluate auditory perception. In this study, we recorded frequency following responses (FFRs) from 13 normal-hearing adult participants to six short music and speech stimuli sharing similar fundamental frequencies but varying in overall spectral and temporal characteristics. Each participant completed a perceptual identification test using the same stimuli. We used linear discriminant analysis to classify FFRs. Results showed statistically significant FFR classification accuracies using both the full response epoch in the time domain (72.3% accuracy, p < 0.001) as well as real and imaginary Fourier coefficients up to 1 kHz (74.6%, p < 0.001). We classified decomposed versions of the responses in order to examine which response features contributed to successful decoding. Classifier accuracies using Fourier magnitude and phase alone in the same frequency range were lower but still significant (58.2% and 41.3% respectively, p < 0.001). Classification of overlapping 20-msec subsets of the FFR in the time domain similarly produced reduced but significant accuracies (42.3%-62.8%, p < 0.001). Participants mean perceptual responses were most accurate (90.6%, p < 0.001). Confusion matrices from FFR classifications and perceptual responses were converted to distance matrices and visualized as dendrograms. FFR classifications and perceptual responses demonstrate similar patterns of confusion across the stimuli. Our results demonstrate that classification can differentiate auditory stimuli from FFR responses with high accuracy. Moreover, the reduced accuracies obtained when the FFR is decomposed in the time and frequency domains suggest that different response features contribute complementary information, similar to how the human auditory system is thought to rely on both timing and frequency information to accurately process sound. Taken together, these results suggest that FFR classification is a promising approach for objective assessment of auditory perception.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- General auditory and speech-specific contributions to cortical envelope tracking revealed using auditory chimeras 97%
- Does amplitude compression help or hinder attentional neural speech tracking? 97%
- Direct cochlear recordings in humans show a theta rhythmic modulation of the auditory nerve by selective attention 97%
Similar papers in this journal
- Attention Decoding at the Cocktail Party: Preserved in Hearing Aid Users, Reduced in Cochlear Implant Users 97%
- Decoding of selective attention to continuous speech from the human auditory brainstem response 97%
- The integration of continuous audio and visual speech in a cocktail-party environment depends on attention 97%
Similar papers in this journal
- Effects of stimulus rate and periodicity on auditory cortical entrainment to continuous sounds 97%
- Individual Differences in Cognition and Perception Predict Neural Processing of Speech in Noise for Audiometrically Normal Listeners. 97%
- Individualized Assays of Temporal Coding in the Ascending Human Auditory System 97%
Similar papers in this journal
Similar papers in this journal
- Cortical compensation for hearing loss, but not age, in neural tracking of the fundamental frequency of the voice 95%
- Mutual Information Analysis of Neural Representations of Speech in Noise in the Aging Midbrain 95%
- Dynamic neural reconstructions of attended object location and features using EEG 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.