Prosodic categories in speech are acoustically multidimensional: evidence from dimension-based statistical learning
Jasmin, K.; Tierney, A. T.; Holt, L.
Show abstract
Segmental speech units such as phonemes are cued by multiple acoustic dimensions (e.g. F0 and duration), but dimensions do not carry equal perceptual weight. The relative perceptual weights of acoustic speech dimensions are not fixed but vary with context. For example, when speech is altered to create an accent in which two acoustic dimensions are correlated in a manner opposite that of long-term experience, the dimension that carries less perceptual weight is down-weighted to contribute less in category decisions. It remains unclear, however, whether this short-term reweighting is limited to segmental categorization, or if it extends to categorization of suprasegmental features which span multiple phonemes, syllables, or words, which would suggest that such "dimension-based statistical learning" is a widespread phenomenon in speech perception. Here we investigated the relative contribution of two acoustic dimensions to word emphasis. Participants categorized instances of a two-word phrase pronounced with typical covariation of fundamental frequency (F0) and duration, and in the context of an artificial accent in which F0 and duration (established in prior research on English speech as primary and secondary dimensions, respectively) covaried atypically. When categorizing accented speech, listeners rapidly down-weighted the secondary dimension (duration) while continuing to rely on the primary dimension (F0). This result indicates that listeners continually track short-term regularities across speech input and dynamically adjust the weight of acoustic evidence for suprasegmental categories. Thus, dimension-based statistical learning appears to be a widespread phenomenon in speech perception extending to both segmental and suprasegmental categorization.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Representations of fricatives in sub-cortical model responses: comparisons with human consonant perception 95%
- Gender and speech material effects on the long-term average speech spectrum, including at extended high frequencies 95%
- Modulation masking and fine structure shape neural envelope coding to predict speech intelligibility across diverse listening conditions 94%
Similar papers in this journal
- Speech auditory-motor adaptation lacks an explicit component: reduced adaptation in adults who stutter reflects limitations in implicit sensorimotor learning. 96%
- Enhanced mismatch negativity in harmonic compared to inharmonic sound sequences 95%
- Perceived multisensory common cause relations shape the ventriloquism effect but only marginally the trial-wise aftereffect 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.