Playback speech detection based on magnitude–phase spectrum

Playback speech detection based on magnitude–phase spectrum
复制标题

DOI:
10.1049/el.2018.0739
复制
发表时间:
2018-06
影响因子:
1.1
通讯作者:
Jichen Yang;Lei-an Liu
Jichen Yang;Lei-an Liu
中科院分区:
工程技术4区
文献类型:
--
作者:
Jichen Yang;Lei-an Liu

文献摘要

被引文献

相似文献

在现有的回放语音检测(PSD)信中,常用的特征往往是从幅度谱中提取的,而没有利用相位谱信息。为了提取更多的PSD鉴别信息,提出了幅度相位谱的思想。在此基础上,提出了一种新的特征,即恒Q幅相倍频程系数(CMPOC)。在ASVspoof 2017评测集上的实验结果表明:(i)CMPOC的性能优于幅度谱和MPS特征。(ii)CMPOC的性能优于一些常用的功能。(iii)他们的系统比一些已知的系统提供更好的性能。
In current playback speech detection (PSD) Letter, commonly used features are often extracted from magnitude spectrum while phase spectrum information is not used. In order to extract more discriminative information for PSD, the idea of magnitude–phase spectrum (MPS) is proposed. Then a new feature based on MPS is proposed, namely constant-Q magnitude–phase octave coefficients (CMPOC). The experimental result on ASVspoof 2017 evaluation set using CMPOC indicates that: (i) the performance of CMPOC is better than features extracted not only from magnitude spectrum but also from MPS. (ii) CMPOC performs better than some commonly used features. (iii) Their system gives better performance than some known systems.