A sinusoidal model based on frequency-to-instantaneous frequency mapping

A sinusoidal model based on frequency-to-instantaneous frequency mapping
复制标题

基于频率到瞬时频率映射的正弦模型

DOI:
10.21437/icslp.2000-906
复制
发表时间:
2000
期刊:
--
影响因子:
--
通讯作者:
Hideki Kawahara
Hideki Kawahara
中科院分区:
--
文献类型:
--
作者:
Parham Zolfaghari;Hideki Kawahara

文献摘要

被引文献

相似文献

本文描述了一种正弦分析与合成框架,该框架采用了一种提取正弦分量和基频的新方法。该方法基于从线性间隔滤波器中心频率到滤波器输出的瞬时频率的映射。通过这种映射得到频域不动点,从而提取输入信号的组成正弦分量。基于该模型的小波表示的鲁棒基频提取技术也被使用。这些构成了正弦分析框架的基本部分,该框架还包括正弦分量轨迹延续方案。为了重建光谱,在合成[3]时采用了逆FFT方法。该模型已被证明可以产生高质量的语音,也适用于其他声源。
In this paper we describe a sinusoidal analysis and synthesis framework which uses a novel method of extracting the sinusoidal components and fundamental frequency. This method is based on a mapping from linearly spaced filter centre frequencies to the instantaneous frequencies of the filter outputs. Frequency domain fixed points are obtained from this mapping which result in the extraction of the constituent sinusoidal components of the input signal. A robust fundamental frequency extraction technique based on a wavelet representation of this model is also used. These form the essential parts of the sinusoidal analysis framework which also includes a sinusoidal component trajectory continuation scheme. In order to reconstruct the spectrum, the inverse FFT method is used in synthesis [3]. This model has been shown to produce speech of high quality and is also applicable to other sound sources.