Temporally variable multi-aspect auditory morphing enabling extrapolation without objective and perceptual breakdown

Temporally variable multi-aspect auditory morphing enabling extrapolation without objective and perceptual breakdown
复制标题

DOI:
10.1109/icassp.2009.4960481
复制
发表时间:
2009-04
期刊:
2009 IEEE International Conference on Acoustics, Speech and Signal Processing
影响因子:
--
通讯作者:
Hideki Kawahara;R. Nisimura;T. Irino;M. Morise;Toru Takahashi;Hideki Banno
Hideki Kawahara;R. Nisimura;T. Irino;M. Morise;Toru Takahashi;Hideki Banno
中科院分区:
其他
文献类型:
--
作者:
Hideki Kawahara;R. Nisimura;T. Irino;M. Morise;Toru Takahashi;Hideki Banno

文献摘要

被引文献

相似文献

提出了一个基于语音分析、修改和再合成系统STRAIGHT的听觉变形通用框架,该框架使代表性方面的每个变形率成为时间的函数,包括时间轴本身。两种类型的算法推导出:一个增量算法的实时操纵变形率和离线后生产应用程序的批处理算法。通过在对数域中的映射函数的导数方面定义变形,消除了在前一制剂中发现的外推情况下的变形再合成的故障。还介绍了一种方法,以减轻感知缺陷外推。
A generalized framework of auditory morphing based on the speech analysis, modification and resynthesis system STRAIGHT is proposed that enables each morphing rate of representational aspects to be a function of time, including the temporal axis itself. Two types of algorithms were derived: an incremental algorithm for real-time manipulation of morphing rates and a batch processing algorithm for off-line post-production applications. By defining morphing in terms of the derivative of mapping functions in the logarithmic domain, breakdown of morphing resynthesis found in the previous formulation in the case of extrapolations was eliminated. A method to alleviate perceptual defects in extrapolation is also introduced.