Towards automatic transcription of spontaneous presentations
Towards automatic transcription of spontaneous presentations
复制标题
实现自发演示的自动转录
DOI:
10.21437/eurospeech.2001-129
复制
发表时间:
2001
期刊:
影响因子:
--
通讯作者:
S. Furui
中科院分区:
文献类型:
--
作者:
T. Shinozaki;Chiori Hori;S. Furui
This paper reports various investigations on recognizing spontaneous presentation speech in connection with the “Spontaneous Speech” national project started in 1999. Presentation speech uttered by 10 male speakers of approximately 4.5 hours duration has been recognized. Experimental results show that acoustic and language modeling based on an actual spontaneous speech corpus is far more effective than conventional modeling based on read speech. The recognition accuracy has a wide speaker-tospeaker variability according to the speaking rate, the number of fillers, the number of repairs, etc. It was confirmed that unsupervised speaker adaptation of acoustic models was effective to improve the recognition accuracy. The recognition accuracy for spontaneous speech is, however, still rather low, and there remains a large number of research issues.