课题基金 / 基金详情

A study on signal extraction in noisy and reverberant environment

A study on signal extraction in noisy and reverberant environment
噪声混响环境下信号提取的研究
批准号:
10680374
负责人:
AKAGI Masato
金额:
$1.98万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (C)
财政年份:
1998
资助国家:
日本
项目状态:
已结题
起止时间:
1998 至 2000

项目摘要

项目成果

AKAGI Masato的其他基金

相似基金

相关文献

中文摘要
翻译
本研究在人类心理声学和听觉生理学知识的基础上,探讨了语音增强和语音分离的模型。利用消去模型对语音进行增强。特别注意使用空间滤波技术降低噪声,并使用频率滤波技术提高基频估计的鲁棒性。这两种技术都采用了消去模型的概念。此外,利用Bregman提出的启发式规则约束,克服了分离两个声源的问题。仿真结果表明,空间滤波和频率滤波都能有效增强语音。因此,这些滤波方法可以有效地用于自动语音识别系统的前端,以及语音特征提取。声音分离模型可以精确地从噪声信号中提取出期望的信号,即使是在波形中。此外,本研究还探讨了基于哺乳动物听觉生理数据的声源方向估计模型。该模型可以解释神经放电传递时间和相位信息与耳际时差准确性之间的关系。
英文摘要
This research discusses models of speech enhancement and segregation based on knowledge about human psychoacoustics and auditory physiology. The cancellation model is used for enhancing speech. Special attention is paid to reducing noise by using a spatial filtering technique, and increasing the robustness of fundamental frequency estimation by using a frequency filtering technique. Both techniques adopt concepts of the cancellation model. In addition, some constraints related to the heuristic regularities proposed by Bregman are used to overcome the problem associated with segregating two acoustic sources. Simulated results show that both spatial and frequency filtering are useful in enhancing speech. As a result, these filtering methods can be used effectively at the front-end of automatic speech recognition systems, and for speech feature extraction. The sound segregation model can precisely extract a desired signal from a noisy signal even in waveforms.Additionally, this research discusses models of sound source direction estimation based on physiological data of mammal audition. The model can explain the relationship between transmission of temporal and phase information by nerve firing and accuracy of interaural time differences.
期刊论文(32)
专著(0)
科研奖励(0)
会议论文
Ito, K.and Akagi, M.: "A study on temporal information based on the synchronization index using a computational model"Proc.WESTPRAC7. 263-266 (2000)
Ito, K. 和 Akagi, M.:“使用计算模型对基于同步索引的时间信息进行研究”Proc.WESTPRAC7。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
Unoki,M.and Akagi,M.: "A method of signal extraction from noisy signal based on auditory scene analysis"Speech Communication. 27,3-4. 261-279 (1999)
Unoki,M. 和 Akagi,M.:“基于听觉场景分析的噪声信号提取信号的方法”语音通信。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
鵜木,赤木: "聴覚の情景解析に基づいた雑音下の調波復合音の一抽出法"電子情報通信学会論文誌. J82-A,10. 1497-1507 (1999)
Uoki,Akagi:“基于听觉场景分析的噪声下谐波解耦声音的提取方法”,电子、信息和通信工程师学会学报 J82-A,10. 1497-1507 (1999)。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
Akagi,Mizumachi Ishimoto and Unokl: "Speech enhancement and segregation based on human auditory mechanims"Proc.IS2000,Aizu. 246-253 (2000)
Akagi、Mizumachi Ishimoto 和 Unokl:“基于人类听觉机制的语音增强和分离”Proc.IS2000,会津。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
32
    A study on new strategy of emotion recognition in speech
    A study on measurement of brain activities with speech production and perception under transferred auditory feedback conditions
    A study on interaction between production and perception in speech communication
    A study on fluctuation of auditory information based on acoustic information deviations and its perception
    海外基金