Neural speech restoration at the cocktail party: Auditory cortex recovers masked speech of both attended and ignored speakers.

Neural speech restoration at the cocktail party: Auditory cortex recovers masked speech of both attended and ignored speakers.
复制标题

DOI:
10.1371/journal.pbio.3000883
复制
发表时间:
2020-10
期刊:
影响因子:
9.8
通讯作者:
Simon JZ
Simon JZ
中科院分区:
生物学1区
文献类型:
--
作者:
Brodbeck C;Jiao A;Hong LE;Simon JZ

文献摘要

参考文献

被引文献

相似文献

人类非常擅长从几个语音源的声音混合中听一个说话者说话。即使没有双耳提示,两个说话者也很容易被分离,但这种能力背后的神经机制还没有得到很好的理解。一种可能性是,早期皮层处理执行声混合的频谱时间分解,允许通过最佳加权重组,折扣源严重重叠的频谱时间区域,重建出席的讲话。使用人类脑磁图(MEG)的反应,2-谈话者的混合物,我们显示了另一种可能性的证据,在早期,积极的隔离发生,即使是强烈的spectrotemporally重叠区域。早期(约70毫秒)反应不重叠spectrotemporal功能被认为是两个谈话者。当相互竞争的说话者的频谱时间特征相互掩盖时,个体的表征会持续存在,但它们会有大约20毫秒的延迟。这表明,听觉皮层恢复被掩盖在混合物中的声学特征,即使它们发生在被忽略的语音中。存在这样的噪声鲁棒的皮质表示,出席以及被忽略的语音中存在的功能,表明一个积极的皮质流分离过程,这可以解释一系列的行为影响被忽略的背景语音。当几个人在说话时,人们如何专注于一个说话者?MEG反应连续两个说话者的混合物表明,即使听众只注意到其中一个说话者,他们的听觉皮层跟踪两个扬声器的声学特征。即使这些特征被其他说话者局部掩蔽,也会发生这种情况。
Humans are remarkably skilled at listening to one speaker out of an acoustic mixture of several speech sources. Two speakers are easily segregated, even without binaural cues, but the neural mechanisms underlying this ability are not well understood. One possibility is that early cortical processing performs a spectrotemporal decomposition of the acoustic mixture, allowing the attended speech to be reconstructed via optimally weighted recombinations that discount spectrotemporal regions where sources heavily overlap. Using human magnetoencephalography (MEG) responses to a 2-talker mixture, we show evidence for an alternative possibility, in which early, active segregation occurs even for strongly spectrotemporally overlapping regions. Early (approximately 70-millisecond) responses to nonoverlapping spectrotemporal features are seen for both talkers. When competing talkers’ spectrotemporal features mask each other, the individual representations persist, but they occur with an approximately 20-millisecond delay. This suggests that the auditory cortex recovers acoustic features that are masked in the mixture, even if they occurred in the ignored speech. The existence of such noise-robust cortical representations, of features present in attended as well as ignored speech, suggests an active cortical stream segregation process, which could explain a range of behavioral effects of ignored background speech. How do humans focus on one speaker when several are talking? MEG responses to a continuous two-talker mixture suggest that, even though listeners attend only to one of the talkers, their auditory cortex tracks acoustic features from both speakers. This occurs even when those features are locally masked by the other speaker.
DOI: 10.1523/jneurosci.5297-12.2013
发表时间: 2013-03-27
期刊: The Journal of neuroscience : the official journal of the Society for Neuroscience
影响因子: --
作者:
Ding N;Simon JZ
通讯作者: Simon JZ
DOI: 10.1121/1.1907229
发表时间: 1953-01-01
影响因子: 2.4
作者:
CHERRY, EC
通讯作者: CHERRY, EC
DOI: 10.1044/1059-0889(2002/004
发表时间: 2002-06-01
影响因子: 1.8
作者:
Burkard, Robert F;Sims, Donald
通讯作者: Sims, Donald
DOI: 10.3758/bf03213894
发表时间: 1994-08-01
期刊: PERCEPTION & PSYCHOPHYSICS
影响因子: --
作者:
BREGMAN, AS;AHAD, P;MELNERICH, L
通讯作者: MELNERICH, L
DOI: 10.1007/s10162-013-0415-y
发表时间: 2013-12-01
影响因子: 2.4
作者:
Billings, Curtis J.;McMillan, Garnett P.;Gille, Sun Mi
通讯作者: Gille, Sun Mi