Information fusion and decision cascading for audio-visual speaker recognition based on time-varying stream reliability prediction
Information fusion and decision cascading for audio-visual speaker recognition based on time-varying stream reliability prediction
复制标题
基于时变流可靠性预测的视听说话人识别信息融合与决策级联
DOI:
10.1109/icme.2003.1221235
复制
发表时间:
2003
期刊:
影响因子:
--
通讯作者:
C. Neti
中科院分区:
文献类型:
--
作者:
U. Chaudhari;G. Ramaswamy;G. Potamianos;C. Neti
We examine the techniques for multi-modal biometric information fusion for verification and identification of speakers, where the reliability of each data stream, either audio of video, is modeled with parameters that are time-varying and depend on the context created by its local behavior. The complementary nature and the time dependent relative reliability of audio and video data is studied in the context of verification and identification, on data collected during a user's interaction with an automated system. Of significance is that this data is not corrupted artificially. Particular focus is directed to verification and its ability to refine identification decisions, by indicating a level of confidence in the system decisions. Results show more striking effects for verification, when using time-dependent fusion, than for identification.
DOI:
10.1109/89.365379
发表时间:
1995-01-01
期刊:
IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
影响因子:
--
作者:
REYNOLDS, DA;ROSE, RC
通讯作者:
ROSE, RC