A fixed dimension and perceptually based dynamic sinusoidal model of speech

A fixed dimension and perceptually based dynamic sinusoidal model of speech
复制标题

固定维度和基于感知的动态正弦语音模型

DOI:
10.1109/icassp.2014.6854810
复制
发表时间:
2014
期刊:
--
影响因子:
--
通讯作者:
Hu Q
Hu Q
中科院分区:
--
文献类型:
--
作者:
Hu Q

文献摘要

参考文献

被引文献

相似文献

本文提出了一种固定的和低维的,基于感知的动态正弦语音模型称为PDM(感知动态模型)。为了减少和固定标准正弦模型中通常使用的正弦分量的数量,我们建议每个临界频带仅使用一个动态正弦分量。对于每个频带,选择具有最大频谱幅度的正弦曲线,并将其与该临界频带的中心频率相关联。该模型在低频下通过在相应频带的边界处并入正弦曲线来扩展,而在较高频率下使用调制噪声分量。进行听力测试,以比较与PDM和国家的最先进的语音模型,其中所有的模型被约束为使用相同数量的参数重建的语音。结果表明,PDM在质量方面明显优于其他系统。
This paper presents a fixed- and low-dimensional, perceptually based dynamic sinusoidal model of speech referred to as PDM (Perceptual Dynamic Model). To decrease and fix the number of sinusoidal components typically used in the standard sinusoidal model, we propose to use only one dynamic sinusoidal component per critical band. For each band, the sinusoid with the maximum spectral amplitude is selected and associated with the centre frequency of that critical band. The model is expanded at low frequencies by incorporating sinusoids at the boundaries of the corresponding bands while at the higher frequencies a modulated noise component is used. A listening test is conducted to compare speech reconstructed with PDM and state-of-the-art models of speech, where all models are constrained to use an equal number of parameters. The results show that PDM is clearly preferred in terms of quality over the other systems.
使用自适应全带谐波模型进行语音分析和合成
DOI: 10.1109/tasl.2013.2266772
发表时间: 2013
期刊: IEEE Transactions on Audio, Speech, and Language Processing
影响因子: --
作者:
G. Degottex;Y. Stylianou
通讯作者: Y. Stylianou
DOI: 10.1109/35.256878
发表时间: 1993-11-01
影响因子: 11.2
作者:
NOLL, P
通讯作者: NOLL, P
Nitech 基于 HMM 的暴雪挑战赛语音合成系统 2005 的详细信息(HMM Speech Synthesis System for Blizzard Challenge 2005)
DOI: --
发表时间: 2006
期刊: 電子情報通信学会技術研究報告 (発表予定)
影响因子: --
作者:
Heiga Zen;Tonioki Toda;Masaru Nakamura;Keiichi Tokuda(全炳河;戸田智基;中村勝;徳田恵一)
通讯作者: 徳田恵一)
语音的全频带自适应谐波表示
DOI: 10.21437/interspeech.2012-138
发表时间: 2012
期刊: IEEE Transactions on Audio, Speech, and Language Processing
影响因子: --
作者:
G. Degottex;Y. Stylianou
通讯作者: Y. Stylianou
音频正弦建模参数选择的自上而下策略
DOI: 10.1109/icassp.2010.5495954
发表时间: 2010
期刊: 2010 IEEE International Conference on Acoustics, Speech and Signal Processing
影响因子: --
作者:
T. Hirvonen;A. Mouchtaris
通讯作者: A. Mouchtaris