SPEECH SYNTHESIS BASED ON SINUSOIDAL MODELING
SPEECH SYNTHESIS BASED ON SINUSOIDAL MODELING
复制标题
基于正弦建模的语音合成
DOI:
--
复制
发表时间:
2004
期刊:
影响因子:
--
通讯作者:
K. Kumar
中科院分区:
文献类型:
--
作者:
K. Kumar
This report presents an introduction to speech synthesis with a brief overview of some methods and their associated problems. The usage of sinusoidal representation of speech waveform for producing synthetic speech is discussed in detail. The sinusoidal model is based on extracting the amplitudes, frequencies, and phases of the component sine waves from the short-time Fourier transform and using them for the production of synthetic speech. The report is based on some of the works carried out on speech modification to achieve desirable characteristics in speech using pitch scaling, time-scaling, and spectral warping. Obtaining expressions like anger, emotion, joy etc., in synthetic speech is one of the major areas of concern these days. This aspect of bringing expression into speech using the sinusoidal model is presented in brief.