Automatic transcription of spontaneous lecture speech
Automatic transcription of spontaneous lecture speech
复制标题
自动转录自发演讲稿
DOI:
10.1109/asru.2001.1034618
复制
发表时间:
2001
期刊:
影响因子:
--
通讯作者:
S. Furui
中科院分区:
文献类型:
--
作者:
Tatsuya Kawahara;H. Nanjo;S. Furui
We introduce our extensive projects on spontaneous speech processing and current trials of lecture speech recognition. A large corpus of lecture presentations and talks is being collected in the project. We have trained initial baseline models and confirmed significant difference of real lectures and written notes. In spontaneous lecture speech, the speaking rate is generally faster and changes a lot, which makes it harder to apply fixed segmentation and decoding settings. Therefore, we propose sequential decoding and speaking-rate dependent decoding strategies. The sequential decoder simultaneously performs automatic segmentation and decoding of input utterances. Then, the most adequate acoustic analysis, phone models and decoding parameters are applied according to the current speaking rate. These strategies achieve improvement on automatic transcription of real lecture speech.
DOI:
10.21437/eurospeech.2001-396
发表时间:
2001-09
期刊:
--
影响因子:
--
作者:
Akinobu Lee;Tatsuya Kawahara;K. Shikano
通讯作者:
Akinobu Lee;Tatsuya Kawahara;K. Shikano