Voice liveness detection algorithms based on pop noise caused by human breath for automatic speaker verification
Voice liveness detection algorithms based on pop noise caused by human breath for automatic speaker verification
复制标题
DOI:
10.21437/interspeech.2015-92
复制
发表时间:
2015-09
期刊:
影响因子:
--
通讯作者:
Sayaka Shiota;F. Villavicencio;J. Yamagishi;Nobutaka Ono;I. Echizen;T. Matsui
中科院分区:
文献类型:
--
作者:
Sayaka Shiota;F. Villavicencio;J. Yamagishi;Nobutaka Ono;I. Echizen;T. Matsui
. Abstract This paper proposes a novel countermeasure framework to detect spoofing attacks to reduce the vulnerability of automatic speaker verification (ASV) systems. Recently, ASV systems have reached equivalent performances equivalent to those of other biometric modalities. However, spoofing techniques against these systems have also progressed drastically. Experimentation using advanced speech synthesis and voice conversion techniques has showed unacceptable false acceptance rates and several new countermeasure algorithms have been explored to detect spoofing materials accurately. However, the counter-measures proposed so far are based on the acoustic differences between natural speech signals and artificial speech signals, expected to become gradually smaller in the near future. In this paper, we focus on voice liveness detection, which aims to validate whether the presented speech signals originated from a live human. We use the phenomenon of pop noise, which is a distortion that happens when human breath reaches a microphone, as liveness evidence. This paper proposes pop noise detection algorithms and shows through an experimental study that they can be used to discriminate live voice signals from artificial ones generated by means of speech synthesis techniques.