Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 Neurophysiological indices of audiovisual speech processing reveal a hierarchy of multisensory integration effects.

Literature DB >> 33824190

Neurophysiological indices of audiovisual speech processing reveal a hierarchy of multisensory integration effects.

Aisling E O'Sullivan¹, Michael J Crosse², Giovanni M Di Liberto³, Alain de Cheveigné^3,4, Edmund C Lalor^5,6.

Abstract

Seeing a speaker's face benefits speech comprehension, especially in challenging listening conditions. This perceptual benefit is thought to stem from the neural integration of visual and auditory speech at multiple stages of processing, whereby movement of a speaker's face provides temporal cues to auditory cortex, and articulatory information from the speaker's mouth can aid recognizing specific linguistic units (e.g., phonemes, syllables). However it remains unclear how the integration of these cues varies as a function of listening conditions. Here we sought to provide insight on these questions by examining EEG responses in humans (males and females) to natural audiovisual, audio, and visual speech in quiet and in noise. We represented our speech stimuli in terms of their spectrograms and their phonetic features, and then quantified the strength of the encoding of those features in the EEG using canonical correlation analysis. The encoding of both spectrotemporal and phonetic features was shown to be more robust in audiovisual speech responses then what would have been expected from the summation of the audio and visual speech responses, suggesting that multisensory integration occurs at both spectrotemporal and phonetic stages of speech processing. We also found evidence to suggest that the integration effects may change with listening conditions, however this was an exploratory analysis and future work will be required to examine this effect using a within-subject design. These findings demonstrate that integration of audio and visual speech occurs at multiple stages along the speech processing hierarchy.SIGNIFICANCE STATEMENT During conversation, visual cues impact our perception of speech. Integration of auditory and visual speech is thought to occur at multiple stages of speech processing and vary flexibly depending on the listening conditions. Here we examine audiovisual integration at two stages of speech processing using the speech spectrogram and a phonetic representation, and test how audiovisual integration adapts to degraded listening conditions. We find significant integration at both of these stages regardless of listening conditions. These findings reveal neural indices of multisensory interactions at different stages of processing and provide support for the multistage integration framework.

Entities: Species

Year: 2021 PMID： 33824190 DOI： 10.1523/JNEUROSCI.0906-20.2021

Source DB: PubMed Journal: J Neurosci ISSN： 0270-6474 Impact factor: 6.167

Keyword Cloud
Cited

7 in total

1. Binding the Acoustic Features of an Auditory Source through Temporal Coherence.

Authors: Mohsen Rezaeizadeh; Shihab Shamma
Journal: Cereb Cortex Commun Date: 2021-10-06

Review 2. Linear Modeling of Neurophysiological Responses to Speech and Other Continuous Stimuli: Methodological Considerations for Applied Research.

Authors: Michael J Crosse; Nathaniel J Zuk; Giovanni M Di Liberto; Aaron R Nidiffer; Sophie Molholm; Edmund C Lalor
Journal: Front Neurosci Date: 2021-11-22 Impact factor: 4.677

Neurophysiological indices of audiovisual speech processing reveal a hierarchy of multisensory integration effects.

1. Binding the Acoustic Features of an Auditory Source through Temporal Coherence.

Review 2. Linear Modeling of Neurophysiological Responses to Speech and Other Continuous Stimuli: Methodological Considerations for Applied Research.

3. Editorial: Neural Tracking: Closing the Gap Between Neurophysiology and Translational Medicine.

4. Intelligibility of audiovisual sentences drives multivoxel response patterns in human superior temporal cortex.

5. Enhancement of speech-in-noise comprehension through vibrotactile stimulation at the syllabic rate.

6. Neurosensory development of the four brainstem-projecting sensory systems and their integration in the telencephalon.

7. Speech-Driven Facial Animations Improve Speech-in-Noise Comprehension of Humans.