课题基金 / 基金详情

Multi-Modal Blind Source Separation for Robot Audition

Multi-Modal Blind Source Separation for Robot Audition
机器人试镜的多模态盲源分离
批准号:
EP/H012842/1
负责人:
Wenwu Wang
金额:
$14.69万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2009
资助国家:
英国
项目状态:
已结题
起止时间:
2009 至 --

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
This proposal draws on expertise in blind source separation and multimodal (audio-visual) speech processing within the Centre for Vision Speech and Signal Processing at University of Surrey. The objective is to perform source separation of the target speech in the presence of multiple competing sound sources in room environments and thereby ultimately provide progress towards automatic machine perception of auditory scenes within an un-controlled natural environment. The fundamental novelty in this work is to exploit visual cues for enhancing the operation of frequency domain blind source separation algorithms. Exploitation of such audio-visual processing is targeted at mitigating the permutation problem, the underdetermined problem (i.e. when the number of sources is greater than the number of microphones), and the reverberation problem, which currently limits the practical applicability of blind source separation algorithms. The focus of the work is therefore on the signal processing algorithms and software tools that can be used to perform automatic separation of sound signals, e.g., for a robot. The body of work in this proposal is underpinned by the substantial experience of the investigators, two from the areas of blind source separation and digital speech processing, and one from the area of computer vision and pattern recognition. The outcomes of the proposed research will be of considerable value to the UK defence industry working especially in the areas of target separation, detection and multi-path mitigation (or dereverberation), with applications in, for example, human-robot interaction, security surveillance and human-computer interaction.
期刊论文(10)
专著(0)
科研奖励(0)
会议论文
DOI: 10.21437/interspeech.2010-192
发表时间: 2010-09
期刊:
影响因子: --
作者: [Qingju Liu;Wenwu Wang;P. Jackson]
通讯作者: Qingju Liu;Wenwu Wang;P. Jackson
DOI: 10.1109/taslp.2014.2320637
发表时间: 2014-09
期刊: IEEE/ACM Transactions on Audio, Speech, and Language Processing
影响因子: --
作者: [Atiyeh Alinaghi;P. Jackson;Qingju Liu;Wenwu Wang]
通讯作者: Atiyeh Alinaghi;P. Jackson;Qingju Liu;Wenwu Wang
Interference Reduction in Reverberant <newline/>Speech Separation With Visual <newline/>Voice Activity Detection
通过视觉 <newline/> 语音活动检测来减少混响 <newline/> 语音分离中的干扰
DOI: 10.1109/tmm.2014.2322824
发表时间: 2014
期刊: IEEE Transactions on Multimedia
影响因子: 7.3
作者: [Liu Q]
通讯作者: Liu Q
DOI: 10.1109/tsp.2013.2277834
发表时间: 2013-11
期刊: IEEE Transactions on Signal Processing
影响因子: 5.4
作者: [Qingju Liu;Wenwu Wang;P. Jackson;M. Barnard;J. Kittler;J. Chambers]
通讯作者: Qingju Liu;Wenwu Wang;P. Jackson;M. Barnard;J. Kittler;J. Chambers
海外基金