Speaking-rate dependent decoding and adaptation for spontaneous lecture speech recognition

Speaking-rate dependent decoding and adaptation for spontaneous lecture speech recognition
复制标题

用于自发演讲语音识别的依赖于语速的解码和自适应

DOI:
10.1109/icassp.2002.5743820
复制
发表时间:
2002
期刊:
2002 IEEE International Conference on Acoustics, Speech, and Signal Processing
影响因子:
--
通讯作者:
Tatsuya Kawahara
Tatsuya Kawahara
中科院分区:
--
文献类型:
--
作者:
H. Nanjo;Tatsuya Kawahara

文献摘要

参考文献

被引文献

相似文献

本文研究了大词汇量自发语音识别中的语速问题。在自发演讲中,语速通常很快,并且在演讲中可能会有很大变化。我们还观察到不同的错误倾向,快速和慢速语音段。因此,我们首先提出了一种依赖于说话速率的解码策略,该策略根据说话速率应用最充分的声学分析、电话模型和解码参数。几种方法进行了研究,他们的选择性应用导致精度的提高。我们还提出了在说话人自适应中利用说话人的语速信息,其中不同的自适应模型被设置为快速和缓慢的话语。结果表明,该方法比常规自适应方法更有效。
This paper addresses the problem of speaking rate in large vocabulary spontaneous speech recognition. In spontaneous lecture speech, the speaking rate is generally fast and may vary a lot within a talk. We also observed different error tendencies for fast and slow speech segments. Therefore, we first present a speaking-rate dependent decoding strategy that applies the most adequate acoustic analysis, phone models and decoding parameters according to the speaking rate. Several methods are investigated and their selective application leads to accuracy improvement. We also propose to make use of speaking-rate information in speaker adaptation, in which the different adapted models are set up for fast and slow utterances. It is confirmed that the method is more effective than normal adaptation.
DOI: 10.21437/eurospeech.2001-396
发表时间: 2001-09
期刊: --
影响因子: --
作者:
Akinobu Lee;Tatsuya Kawahara;K. Shikano
通讯作者: Akinobu Lee;Tatsuya Kawahara;K. Shikano
DOI: 10.1006/csla.1995.0010
发表时间: 1995-04-01
影响因子: 4.3
作者:
LEGGETTER, CJ;WOODLAND, PC
通讯作者: WOODLAND, PC