Robust Automatic Human Identification Using Face, Mouth, and Acoustic Information
Robust Automatic Human Identification Using Face, Mouth, and Acoustic Information
复制标题
使用面部、嘴巴和声音信息进行稳健的自动人体识别
DOI:
--
复制
发表时间:
2005
期刊:
影响因子:
--
通讯作者:
R. Reilly
中科院分区:
文献类型:
--
作者:
N. Fox;R. Gross;J. Cohn;R. Reilly
Discriminatory information about person identity is multimodal. Yet, most person recognition systems are unimodal, e.g. the use of facial appearance. With a view to exploiting the complementary nature of different modes of information and increasing pattern recognition robustness to test signal degradation, we developed a multiple expert biometric person identification system that combines information from three experts: face, visual speech, and audio. The system uses multimodal fusion in an automatic unsupervised manner, adapting to the local performance and output reliability of each of the experts. The expert weightings are chosen automatically such that the reliability measure of the combined scores is maximized. To test system robustness to train/test mismatch, we used a broad range of Gaussian noise and JPEG compression to degrade the audio and visual signals, respectively. Experiments were carried out on the XM2VTS database. The multimodal expert system out performed each of the single experts in all comparisons. At severe audio and visual mismatch levels tested, the audio, mouth, face, and tri-expert fusion accuracies were 37.1%, 48%, 75%, and 92.7% respectively, representing a relative improvement of 23.6% over the best performing expert.