The effect of semantic content on the perception of audiovisual movieclips
Kesoglou, A. M.; Mikellidou, K.
Show abstract
Our brain is skilled with the ability to perceive and process multimodal stimuli. This process known as crossmodal perceptual integration, has been in the research spotlight for a long time, providing evidence for the integration of information coming from different modalities. Prior research mostly utilized pictures and focused on the semantic content of a single sound or word. The present study aims to investigate crossmodal perceptual integration in realistic conditions using short movieclips (1500ms) and auditory meaningful three-word sentences to evaluate target detection in terms of accuracy and response times. Two experimental tasks were developed using PsychoPy, where participants had to indicate whether a target (noun for Experiment 1, verb for Experiment 2) was present or absent. In trials without a target, target-related information was always present, either through one of the two senses (vision or audition; incongruent condition) or through both senses (congruent condition). We observed superior performance when the target was absent generally (Mexp1 =93.9%, SDexp1 =0.04, Mexp2 =84.2%, SDexp2 =0.182) compared to when it was present (Mexp1 =82.3%, SDexp1 =0.143, Mexp2 =73.7%, SDexp2 =0.184). Moreover, superior performance was noted in incongruent target-related movieclips, which significantly decreased during congruent target-related movieclips. In Experiment 1, we observed that in the audio condition when the target-related word was a noun, participant performance was superior compared to when it was a verb (M=99.4% vs. M=86.7%; tincVerb vs. incNoun =-8.428, p=.001). In Experiment 2, the judgment scores were similarly high in incongruent movieclips and significantly lower in congruent ones regardless of whether the target-related information presented was a verb or noun. The present results provide evidence regarding the role of complexity of semantics, and especially the diverse role verbs and nouns could play in crossmodal perceptual integration in more realistic situations. Our findings can enrich the content of learning techniques, as well as the design of AI models, by taking advantage of the supporting role of semantic audiovisual information, while taking into consideration the potential confusion that the complexity of semantic information can induce to perception experience.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Let Me Make You Happy, And I'll Tell You How you look around: Using an Approach-Avoidance Task as an Embodied Emotion Prime in a Free-Viewing Task 95%
- Processing Stage Flexibility of the SNARC effect: Task Relevance or Magnitude Relevance? 94%
- Frequency Effects on Spelling Error Recognition: An ERP Study 93%
Similar papers in this journal
- Visualizing sounds: training-induced plasticity with a visual-to-auditory conversion device 97%
- Can you affect me? The influence of vitality forms on action perception and motor response 95%
- Auditory-cognitive determinants of speech-in-noise perception: structural equation modelling of a large sample 95%
Similar papers in this journal
- Time-oriented attention improves accuracy in a paced finger tapping task 95%
- Sounds and Sights in Sequence Learning: Can Accessory Auditory Cues Enhance Motor Task Performance? 95%
- Top-down modulation of neural envelope tracking: the interplay with behavioral, self-reported and neural measures of listening effort 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.