Multiple Pitch Tracking and Harmonic Segregation Algorithm for Auditory Scene Analysis

Multiple Pitch Tracking and Harmonic Segregation Algorithm for Auditory Scene Analysis
复制标题

用于听觉场景分析的多音高跟踪和谐波分离算法

DOI:
10.9746/sicetr1965.34.483
复制
发表时间:
1998
期刊:
Journal of the Society of Instrument and Control Engineers
影响因子:
--
通讯作者:
S. Ando
S. Ando
中科院分区:
--
文献类型:
--
作者:
Kazuki Nishi;M. Abe;S. Ando

文献摘要

被引文献

相似文献

本文提出了一种新的听觉场景分析算法,该算法可以自动地从多个背景声流中选择一个单一的声流,并重建其原始波形。该算法的特点是:1)有效地利用了谐波结构,流的特征在于音调及其动态;因此,它将被时变梳状滤波器分离。2)在小波域中进行分析和合成,以提取流,使其频谱不随基音估计误差而变化。3)使用Parzen密度估计和非参数卡尔曼滤波器对候选音调进行最佳和多模态跟踪。该算法是检查几个模拟和实验,使用一些人工和真实的世界的数据。
This paper describes a new algorithm of the auditory scene analysis, by which we can automatically choose a single sound stream from many background ones and reconstruct its original waveform. The features of our algorithm are as follows: 1) Efficient use of harmonics structure, i.e., stream is characterized by pitch and its dynamics; therefore it will be segregated by a time-varying comb filter. 2) Analysis and synthesis in the wavelet domain for extracting a stream so that its spectrum is invariant with the pitch estimation error. 3) The use of Parzen's density estimate and non-parametric Kalman filter for the optimum and multimodal tracking of pitch candidates. This algorithm is examined by several simulations and experiments using some artificial and real world data.