The effect of prior visual information on recognition of speech and sounds

The effect of prior visual information on recognition of speech and sounds
复制标题

DOI:
10.1093/cercor/bhm091
复制
发表时间:
2008-03-01
期刊:
影响因子:
3.7
通讯作者:
Friston, Karl J.
Friston, Karl J.
中科院分区:
医学2区
文献类型:
--
作者:
Noppeney, Uta;Josephs, Oliver;Friston, Karl J.

文献摘要

被引文献

相似文献

为了识别和分类复杂的刺激,如熟悉的物体或言语,人类的大脑整合了从感官输入中提取的多个层次的信息。通过对口语单词和声音的交叉模态启动,本功能性磁共振成像研究确定了三种不同类型的视听觉不一致效应:1)左侧颞上沟(STS)中的口语单词,2)左侧角回(AG)中的环境声音,以及3)外侧和内侧前额叶皮层(IFS/mPFC)中的单词和声音都具有选择性。从认知的角度来看,这些不一致效应表明,先验视觉信息在多个层面上影响语音和声音识别的神经过程,其中STS参与语音加工,AG参与语义加工,mPFC/IFS参与高级概念加工。在神经机制方面,有效的连通性分析(动态因果模型)表明,这些不一致效应可能通过从早期听觉区域到中间多感觉整合区域(即STS和AG)的更大的自下而上效应出现。这与皮层分层贝叶斯推理的预测编码观点是一致的,其中预测错误的领域(语音与语义)决定了其区域表达(颞中回/STS vs. AG/顶叶内沟)。
To identify and categorize complex stimuli such as familiar objects or speech, the human brain integrates information that is abstracted at multiple levels from its sensory inputs. Using cross-modal priming for spoken words and sounds, this functional magnetic resonance imaging study identified 3 distinct classes of visuoauditory incongruency effects: visuoauditory incongruency effects were selective for 1) spoken words in the left superior temporal sulcus (STS), 2) environmental sounds in the left angular gyrus (AG), and 3) both words and sounds in the lateral and medial prefrontal cortices (IFS/mPFC). From a cognitive perspective, these incongruency effects suggest that prior visual information influences the neural processes underlying speech and sound recognition at multiple levels, with the STS being involved in phonological, AG in semantic, and mPFC/IFS in higher conceptual processing. In terms of neural mechanisms, effective connectivity analyses (dynamic causal modeling) suggest that these incongruency effects may emerge via greater bottom-up effects from early auditory regions to intermediate multisensory integration areas (i.e., STS and AG). This is consistent with a predictive coding perspective on hierarchical Bayesian inference in the cortex where the domain of the prediction error (phonological vs. semantic) determines its regional expression (middle temporal gyrus/STS vs. AG/intraparietal sulcus).