Speaking-rate dependent decoding and adaptation for spontaneous lecture speech recognition
Speaking-rate dependent decoding and adaptation for spontaneous lecture speech recognition
复制标题
用于自发演讲语音识别的依赖于语速的解码和自适应
DOI:
10.1109/icassp.2002.5743820
复制
发表时间:
2002
期刊:
影响因子:
--
通讯作者:
Tatsuya Kawahara
中科院分区:
文献类型:
--
作者:
H. Nanjo;Tatsuya Kawahara
This paper addresses the problem of speaking rate in large vocabulary spontaneous speech recognition. In spontaneous lecture speech, the speaking rate is generally fast and may vary a lot within a talk. We also observed different error tendencies for fast and slow speech segments. Therefore, we first present a speaking-rate dependent decoding strategy that applies the most adequate acoustic analysis, phone models and decoding parameters according to the speaking rate. Several methods are investigated and their selective application leads to accuracy improvement. We also propose to make use of speaking-rate information in speaker adaptation, in which the different adapted models are set up for fast and slow utterances. It is confirmed that the method is more effective than normal adaptation.
DOI:
10.21437/eurospeech.2001-396
发表时间:
2001-09
期刊:
--
影响因子:
--
作者:
Akinobu Lee;Tatsuya Kawahara;K. Shikano
通讯作者:
Akinobu Lee;Tatsuya Kawahara;K. Shikano
影响因子:
4.3
作者:
LEGGETTER, CJ;WOODLAND, PC
通讯作者:
WOODLAND, PC