YIN, a fundamental frequency estimator for speech and music

YIN, a fundamental frequency estimator for speech and music
复制标题

DOI:
10.1121/1.1458024
复制
发表时间:
2002-04-01
影响因子:
2.4
通讯作者:
Kawahara, H
Kawahara, H
中科院分区:
物理与天体物理3区
文献类型:
--
作者:
de Cheveigné, A;Kawahara, H

文献摘要

被引文献

相似文献

本文提出了一种用于估计语音或音乐声音基频(F - 0)的算法。它基于著名的自相关方法,并进行了一些修改以防止错误。该算法具有几个理想的特性。在一个同时记录了喉图信号的语音数据库上进行评估时,其错误率比最佳竞争方法低约三倍。频率搜索范围没有上限,因此该算法适用于高音调的声音和音乐。该算法相对简单,可以高效实现且延迟较低,并且涉及很少需要调整的参数。它基于一种信号模型(周期信号),该模型可以通过几种方式扩展,以处理特定应用中出现的各种形式的非周期性。最后,可以与听觉处理模型进行有趣的类比。(C)2002美国声学学会。
An algorithm is presented for the estimation of the fundamental frequency (F-0) of speech or musical sounds. It is based on the well-known autocorrelation method with a number of modifications that combine to prevent errors. The algorithm has several desirable features. Error rates are about three times lower than the best competing methods, as evaluated over a database of speech recorded together with a laryngograph signal. There is no upper limit on the frequency search range, so the algorithm is suited for high-pitched voices and music. The algorithm is relatively simple and may be implemented efficiently and with low latency, and it involves few parameters that must be tuned. It is based on a signal model (periodic signal) that may be extended in several ways to handle various forms of aperiodicity that occur in particular applications. Finally, interesting parallels may be drawn with models of auditory processing. (C) 2002 Acoustical Society of America.