Music Information Processing Using Continuous Speech Recognition Methods
Music Information Processing Using Continuous Speech Recognition Methods
批准号:
14380156
负责人:
SAGAYAMA Shigeki
金额:
$10.82万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2002
资助国家:
日本
项目状态:
已结题
起止时间:
2002 至 2004
中文摘要
点击翻译按钮获取中文摘要
英文摘要
We formulated music rhythm recognition for ranscribing MIDI data into music score as a Viterbi path search problem in HMM where hidden states and output probabilities represent the intended note values and actually played note lengths, respectively. We also solved rhythm recognition of polyphonic music by reducing polyphony intomonophony. Tempo modeling and tempo change detection were enabled with segmental k-means algorithm for speech recognition.Harmonization (chord finding) of given melodies was formulated as an isomorphic problem as continuous speech recognition by defining output by the given melody, hidden states by the chord behind the melody and stochastic language model by chord sequences. Automatic counterpoint was developed with a two-step maximum likelihood approach consisting of rhythm design and pitch allocation solved by dynamic programming.In polyphonic signal analysis, an algorithm named Harmonic-structured Clustering was developed based on the k-means clustering algorithm under harmonic constraint by modeling the framewise observed spectrum as overlapped harmonic structures and considering that the distributed energy in harmonic structure belongs to a single cluster. Furthermore, by introducing the probabilistic assignment to clusters, k-means was generalized into the EM-algorithm and attained higher performance of multi-pitch estimation. Utilizing an information criterion such as AIC, the number of sources and octave location were also enabled."Specmurt analysis" was proposed for polyphonic signal analysis. The inverse Fourier transform of linear spectrum with log-frequency was called "specmurt". Along log-scaled frequency, observed linear spectrum is regarded as convolution of distribution density of fundamental frequencies and harmonic structures of multiple tones which are assumed identical. This idea opened up a new signal processing capabilities.
期刊论文(223)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Time-Space Clustering for Multi-pitch Spectral Segregation Using Kernel Audio Stream Model
使用内核音频流模型进行多音高频谱分离的时空聚类
DOI:
--
发表时间:
2005
期刊:
The 2005 Spring Meeting of the Acoustic Society of Japan 3-7-19
影响因子:
--
作者:
[Hirokazu Kameoka, Takuya Nishimoto, Shigeki Sagayama]
通讯作者:
Shigeki Sagayama
Extraction of Multiple Fundamental Frequencies from Polyphonic Music Using Harmonic Clustering
使用谐波聚类从复调音乐中提取多个基本频率
DOI:
--
发表时间:
2004
期刊:
Proc. International Congress on Acoustics (ICA) (Kyoto, Japan)
影响因子:
--
作者:
[Hirokazu Kameoka, Takuya Nishimoto, Shigeki Sagayama]
通讯作者:
Shigeki Sagayama
調波スペクトル分離の原理Harmonic Clusteringと赤池情報量規準による多声部楽曲音響信号の同時発音数および多重ピッチの推定
使用谐波聚类、谐波频谱分离原理和 Akaike 信息准则估计复调音乐音频信号的同时语音数量和多个音高
DOI:
--
发表时间:
2004
期刊:
日本音響学会7月音楽音響研究会資料
影响因子:
--
作者:
[亀岡弘和, 西本卓也, 嵯峨山茂樹]
通讯作者:
嵯峨山茂樹
ハーモニック・クラスタリングによる多重音信号音高抽出における音源数とオクターブ位置推定
使用谐波聚类估计多音信号基音提取中的声源数量和八度位置
DOI:
--
发表时间:
2003
期刊:
日本音響学会2003年秋季研究発表会講演論文集
影响因子:
--
作者:
[亀岡 弘和, 西本 卓也, 嵯峨山 茂樹]
通讯作者:
嵯峨山 茂樹
確率モデルによる多声楽曲MIDI演奏からの楽譜推定
使用概率模型根据和弦 MIDI 演奏估计乐谱
DOI:
--
发表时间:
2003
期刊:
情報処理学会研究報告 2003-MUS-50
影响因子:
--
作者:
[武田 晴登, 西本 卓也, 篠田 浩一, 嵯峨山 茂樹]
通讯作者:
嵯峨山 茂樹
共 92 条
Versatile music processing by combining statistical signal processing and music theory
-
批准号:23240021
-
项目类别:Grant-in-Aid for Scientific Research (A)
-
资助金额:$24.79万
-
财政年份:2011
-
负责人:SAGAYAMA Shigeki
-
依托单位:
Analysis, Recognition, Manipulation and Generation of Music Signal and Information based on Mathematical Models
-
批准号:20240017
-
项目类别:Grant-in-Aid for Scientific Research (A)
-
资助金额:$23.96万
-
财政年份:2008
-
负责人:SAGAYAMA Shigeki
-
依托单位:
Research on Signal and Information Processing for Automatic Music Analysis, Recognition and Generation
-
批准号:17300054
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$10.72万
-
财政年份:2005
-
负责人:SAGAYAMA Shigeki
-
依托单位:
Recognition of Cursive/Blind Kanji Handwriting Utilizing the Contuinous Speech Recognition Approach
-
批准号:11480074
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$6.78万
-
财政年份:1999
-
负责人:SAGAYAMA Shigeki
-
依托单位:
海外基金