Computational Models of Auditory Scene Analysis: A Review

Computational Models of Auditory Scene Analysis: A Review
复制标题

听觉场景分析的计算模型:回顾

DOI:
--
复制
发表时间:
2016
影响因子:
4.3
通讯作者:
I. Winkler
I. Winkler
中科院分区:
医学2区
文献类型:
--
作者:
B. T. Szabó;S. Denham;I. Winkler

文献摘要

参考文献

被引文献

相似文献

听觉场景分析(阿萨)是指将复杂的声学输入解析成代表物理源或时间声音模式(例如旋律)的听觉感知对象的过程,其有助于声波到达耳朵。一些新的计算模型占一些感知现象的阿萨最近已经出版。在这里,我们提供了一个理论上的动机审查这些计算模型,旨在将其指导原则的阿萨的理论框架的核心问题。具体来说,我们问他们如何实现分组和分离的声音元素,以及他们是否实现某种形式的竞争之间的替代解释的声音输入。我们认为,在何种程度上,他们包括预测过程,作为重要的当前理论表明,感知是固有的预测,以及他们如何被评估。我们的结论是,目前的计算模型的阿萨是零碎的意义上说,而不是提供一般的竞争解释阿萨,他们专注于评估的实用程序的特定过程(或算法),找到复杂的声学信号的原因。这使得开放的可能性,将互补方面的模型到一个更全面的阿萨理论。
Auditory scene analysis (ASA) refers to the process (es) of parsing the complex acoustic input into auditory perceptual objects representing either physical sources or temporal sound patterns, such as melodies, which contributed to the sound waves reaching the ears. A number of new computational models accounting for some of the perceptual phenomena of ASA have been published recently. Here we provide a theoretically motivated review of these computational models, aiming to relate their guiding principles to the central issues of the theoretical framework of ASA. Specifically, we ask how they achieve the grouping and separation of sound elements and whether they implement some form of competition between alternative interpretations of the sound input. We consider the extent to which they include predictive processes, as important current theories suggest that perception is inherently predictive, and also how they have been evaluated. We conclude that current computational models of ASA are fragmentary in the sense that rather than providing general competing interpretations of ASA, they focus on assessing the utility of specific processes (or algorithms) for finding the causes of the complex acoustic signal. This leaves open the possibility for integrating complementary aspects of the models into a more comprehensive theory of ASA.
DOI: 10.1152/jn.00788.2006
发表时间: 2007-03-01
影响因子: 2.5
作者:
Wilson, E. Courtenay;Melcher, Jennifer R.;Oxenham, Andrew J.
通讯作者: Oxenham, Andrew J.
DOI: 10.1121/1.1836832
发表时间: 2005-02-01
影响因子: 2.4
作者:
Helfer, KS;Freyman, RL
通讯作者: Freyman, RL
时间连贯性和复杂声音的流动。
DOI: 10.1007/978-1-4614-1590-9_59
发表时间: 2013
影响因子: --
作者:
Shamma,Shihab;Elhilali,Mounya;Ma,Ling;Micheyl,Christophe;Oxenham,AndrewJ;Pressnitzer,Daniel;Yin,Pingbo;Xu,Yanbo
通讯作者: Xu,Yanbo
DOI: 10.1016/j.ijpsycho.2014.05.005
发表时间: 2015-02
影响因子: 3
作者:
Simon, Jonathan Z.
通讯作者: Simon, Jonathan Z.
DOI: 10.1121/1.410023
发表时间: 1994-06-01
影响因子: 2.4
作者:
KIDD, G;MASON, CR;COLBURN, HS
通讯作者: COLBURN, HS