Robust sound classification through the representation of similarity using response fields derived from stimuli during early experience

Robust sound classification through the representation of similarity using response fields derived from stimuli during early experience
复制标题

通过使用从早期体验中的刺激衍生的响应场来表示相似性,进行稳健的声音分类

DOI:
10.1007/s00422-005-0560-4
复制
发表时间:
2005
影响因子:
1.9
通讯作者:
S. Denham
S. Denham
中科院分区:
工程技术3区
文献类型:
--
作者:
M. Coath;S. Denham

文献摘要

参考文献

被引文献

相似文献

听觉处理模型,特别是语音模型,面临着许多困难。其中包括说话者之间的可变性,语音速率的可变性,以及对时间压缩等适度失真的鲁棒性。我们构建了一个系统的基础上合奏的特征检测器来自片段的发作敏感的声音表示。该方法基于“光谱-时间响应场”的思想,并使用卷积来测量特征检测器和刺激之间的时间相似度。从合奏的输出被用来获得分割线索和模式的反应,这是用来训练人工神经网络(ANN)分类。这使我们能够估计输入类和输出类之间互信息的下界。我们的研究结果表明,有显着的信息在我们的系统的输出,这是强大的功能集,时间压缩的刺激,和扬声器的变化的确切选择。此外,在刺激中的时间压缩的鲁棒性具有与人类心理物理学共同的特征。使用来自非语音声音片段的特征检测器的类似实验表现不太好。这一结果是有趣的,因为在发育阶段暴露于贫困听觉环境的动物中,皮层发育异常。
Models of auditory processing, particularly of speech, face many difficulties. Included in these are variability among speakers, variability in speech rate, and robustness to moderate distortions such as time compression. We constructed a system based on ensembles of feature detectors derived from fragments of an onset-sensitive sound representation. This method is based on the idea of ‘spectro-temporal response fields’ and uses convolution to measure the degree of similarity through time between the feature detectors and the stimulus. The output from the ensemble was used to derive segmentation cues and patterns of response, which were used to train an artificial neural network (ANN) classifier. This allowed us to estimate a lower bound for the mutual information between the class of the input and the class of the output. Our results suggest that there is significant information in the output of our system, and that this is robust with respect to the exact choice of feature set, time compression in the stimulus, and speaker variation. In addition, the robustness to time compression in the stimulus has features in common with human psychophysics. Similar experiments using feature detectors derived from fragments of non-speech sounds performed less well. This result is interesting in the light of results showing aberrant cortical development in animals exposed to impoverished auditory environments during the developmental phase.
DOI: 10.1152/jn.2001.85.3.1220
发表时间: 2001-03-01
影响因子: 2.5
作者:
Depireux, DA;Simon, JZ;Shamma, SA
通讯作者: Shamma, SA