Information fusion and decision cascading for audio-visual speaker recognition based on time-varying stream reliability prediction

Information fusion and decision cascading for audio-visual speaker recognition based on time-varying stream reliability prediction
复制标题

基于时变流可靠性预测的视听说话人识别信息融合与决策级联

DOI:
10.1109/icme.2003.1221235
复制
发表时间:
2003
期刊:
2003 International Conference on Multimedia and Expo. ICME '03. Proceedings (Cat. No.03TH8698)
影响因子:
--
通讯作者:
C. Neti
C. Neti
中科院分区:
--
文献类型:
--
作者:
U. Chaudhari;G. Ramaswamy;G. Potamianos;C. Neti

文献摘要

参考文献

被引文献

相似文献

我们研究了用于说话人验证和识别的多模式生物特征信息融合技术,其中每个数据流的可靠性,无论是音频还是视频,都是用时变的参数建模的,并且取决于其局部行为所产生的上下文。音频和视频数据的互补性和依赖时间的相对可靠性是在验证和识别的背景下,根据用户与自动化系统交互期间收集的数据来研究的。重要的是,这些数据并没有被人为破坏。特别侧重于核查及其通过表明对系统决策的信任程度来完善识别决策的能力。结果表明,当使用依赖于时间的融合时,验证效果比识别效果更显著。
We examine the techniques for multi-modal biometric information fusion for verification and identification of speakers, where the reliability of each data stream, either audio of video, is modeled with parameters that are time-varying and depend on the context created by its local behavior. The complementary nature and the time dependent relative reliability of audio and video data is studied in the context of verification and identification, on data collected during a user's interaction with an automated system. Of significance is that this data is not corrupted artificially. Particular focus is directed to verification and its ability to refine identification decisions, by indicating a level of confidence in the system decisions. Results show more striking effects for verification, when using time-dependent fusion, than for identification.
DOI: 10.1109/89.365379
发表时间: 1995-01-01
期刊: IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
影响因子: --
作者:
REYNOLDS, DA;ROSE, RC
通讯作者: ROSE, RC