Speaker Verification with Adaptive Spectral Subband Centroids

Speaker Verification with Adaptive Spectral Subband Centroids
复制标题

使用自适应频谱子带质心进行扬声器验证

DOI:
--
复制
发表时间:
2007
期刊:
International Conference on Biometrics
影响因子:
--
通讯作者:
Ye Wang
Ye Wang
中科院分区:
--
文献类型:
--
作者:
T. Kinnunen;Bingjun Zhang;Jia Zhu;Ye Wang

文献摘要

被引文献

相似文献

频谱子带质心(SSC)已被用作语音和说话人识别中倒谱系数的附加特征。ssc被计算为子带的质心频率,它们捕获短期频谱的主导频率。在基线SSC方法中,子带滤波器是预先指定的。为了更好地适应形成峰运动和其他动态现象,我们建议使用全局最优标量量化方案在逐帧的基础上调整子带滤波器边界。该方法只有一个控制参数,即子带数。NIST 2001任务上的说话人验证结果表明,参数的选择并不重要,该方法不需要额外的特征归一化。
Spectral subband centroids (SSC) have been used as an additional feature to cepstral coefficients in speech and speaker recognition. SSCs are computed as the centroid frequencies of subbands and they capture the dominant frequencies of the short-term spectrum. In the baseline SSC method, the subband filters are pre-specified. To allow better adaptation to formant movements and other dynamic phenomena, we propose to adapt the subband filter boundaries on a frame-by-frame basis using a globally optimal scalar quantization scheme. The method has only one control parameter, the number of subbands. Speaker verification results on the NIST 2001 task indicate that the selection of the parameter is not critical and that the method does not require additional feature normalization.