A full-band adaptive harmonic representation of speech
A full-band adaptive harmonic representation of speech
复制标题
语音的全频带自适应谐波表示
DOI:
10.21437/interspeech.2012-138
复制
发表时间:
2012
期刊:
影响因子:
--
通讯作者:
Y. Stylianou
中科院分区:
文献类型:
--
作者:
G. Degottex;Y. Stylianou
In this paper we present a full-band Adaptive Harmonic Model (aHM) that is able to accurately reconstruct stationary and non stationary parts of speech. The model does not require any voiced/unvoiced decision, neither an accurate estimation of the pitch contour. Its robustness is based on the previously suggested adaptive Quasi-Harmonic model (aQHM), which provides a mechanism for frequency correction and adaptivity of its basis functions to the characteristics of the input signal. The suggested method overcomes limitations of the initial method based on aQHM in detecting frequency tracks over time, especially at mid and high frequencies, by employing a bandlimited iterative procedure for the re-estimation of the fundamental frequency. Listening tests show that reconstructed speech using aHM is mainly indistinguishable from the original signal, outperforming standard sinusoidal models (SM) and the aQHMbased method, while it uses less parameters for the reconstruction than SM.