A study on signal extraction in noisy and reverberant environment
A study on signal extraction in noisy and reverberant environment
批准号:
10680374
负责人:
AKAGI Masato
金额:
$1.98万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (C)
财政年份:
1998
资助国家:
日本
项目状态:
已结题
起止时间:
1998 至 2000
中文摘要
本研究以人类心理声学和听觉生理学的知识为基础,探讨语音增强和分离的模型。消除模型用于增强语音。特别注意通过使用空间滤波技术来降低噪声,并且通过使用频率滤波技术来增加基频估计的鲁棒性。这两种技术都采用了取消模型的概念。此外,Bregman提出的启发式搜索的一些约束条件被用来克服与隔离两个声源相关的问题。仿真结果表明,空域滤波和频域滤波都能有效地增强语音。因此,这些滤波方法可以有效地用于自动语音识别系统的前端,并用于语音特征提取。声音分离模型可以从噪声信号中精确地提取出期望信号,即使在波形中也是如此。此外,本研究还讨论了基于哺乳动物听觉生理数据的声源方向估计模型。该模型可以解释神经放电传递的时间和相位信息与耳间时差精度之间的关系。
英文摘要
This research discusses models of speech enhancement and segregation based on knowledge about human psychoacoustics and auditory physiology. The cancellation model is used for enhancing speech. Special attention is paid to reducing noise by using a spatial filtering technique, and increasing the robustness of fundamental frequency estimation by using a frequency filtering technique. Both techniques adopt concepts of the cancellation model. In addition, some constraints related to the heuristic regularities proposed by Bregman are used to overcome the problem associated with segregating two acoustic sources. Simulated results show that both spatial and frequency filtering are useful in enhancing speech. As a result, these filtering methods can be used effectively at the front-end of automatic speech recognition systems, and for speech feature extraction. The sound segregation model can precisely extract a desired signal from a noisy signal even in waveforms.Additionally, this research discusses models of sound source direction estimation based on physiological data of mammal audition. The model can explain the relationship between transmission of temporal and phase information by nerve firing and accuracy of interaural time differences.
期刊论文(32)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Ito, K.and Akagi, M.: "A study on temporal information based on the synchronization index using a computational model"Proc.WESTPRAC7. 263-266 (2000)
Ito, K. 和 Akagi, M.:“使用计算模型对基于同步索引的时间信息进行研究”Proc.WESTPRAC7。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Unoki,M.and Akagi,M.: "A method of signal extraction from noisy signal based on auditory scene analysis"Speech Communication. 27,3-4. 261-279 (1999)
Unoki,M. 和 Akagi,M.:“基于听觉场景分析的噪声信号提取信号的方法”语音通信。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
鵜木,赤木: "聴覚の情景解析に基づいた雑音下の調波復合音の一抽出法"電子情報通信学会論文誌. J82-A,10. 1497-1507 (1999)
Uoki,Akagi:“基于听觉场景分析的噪声下谐波解耦声音的提取方法”,电子、信息和通信工程师学会学报 J82-A,10. 1497-1507 (1999)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Akagi,Mizumachi Ishimoto and Unokl: "Speech enhancement and segregation based on human auditory mechanims"Proc.IS2000,Aizu. 246-253 (2000)
Akagi、Mizumachi Ishimoto 和 Unokl:“基于人类听觉机制的语音增强和分离”Proc.IS2000,会津。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 32 条
A study on new strategy of emotion recognition in speech
-
批准号:22650032
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.12万
-
财政年份:2010
-
负责人:AKAGI Masato
-
依托单位:
A study on measurement of brain activities with speech production and perception under transferred auditory feedback conditions
-
批准号:20300064
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$12.15万
-
财政年份:2008
-
负责人:AKAGI Masato
-
依托单位:
A study on interaction between production and perception in speech communication
-
批准号:16300053
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$8.13万
-
财政年份:2004
-
负责人:AKAGI Masato
-
依托单位:
A study on fluctuation of auditory information based on acoustic information deviations and its perception
-
批准号:13610079
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.11万
-
财政年份:2001
-
负责人:AKAGI Masato
-
依托单位:
Study on Speaker Individuality in Speech and its Control
-
批准号:07680388
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.41万
-
财政年份:1995
-
负责人:AKAGI Masato
-
依托单位:
海外基金