Automatic indexing for lecture speech and its advanced utilization through speech interaction
Automatic indexing for lecture speech and its advanced utilization through speech interaction
批准号:
17300064
负责人:
NAKAGAWA Seiichi
金额:
$10.26万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2005
资助国家:
日本
项目状态:
已结题
起止时间:
2005 至 2007
中文摘要
我们收集了16名发言者、114场讲座、3860分钟的课堂演讲,并发布了语料库。我们开发了我校研究生课堂讲稿数据的自动语音识别、句子提取、切分/标引、语音检索和讲课浏览系统的构建过程。这些过程是必要的,以提高播放声音或视频数据的可用性,在讲课、总结和索引讲稿或视频的情况下,使学生能够更有效地学习。我们的目标是构建这样一个有条理的授课内容框架。为了实现这一目标,我们首先研究了录音方法对语音识别性能的影响。事实证明,高质量的手持麦克风和低质量的翻领麦克风在准确度上存在23%的差异。此外,我们利用相关的Web文本对领域相关的语言模型进行了改进,并提出了一个填充词插入模型。其次,我们尝试通过提取重要句子来进行自动摘要,我们得到了0.319-0.456的κ值,与人类的0.407-0.477不相上下。最后,我们构建了一个讲座浏览系统,使用户能够更有效地利用上述过程的结果进行学习,并对其进行了评估
英文摘要
We collected the class room lecture speech consisting of 16 speakers, 114 lectures, and 3860 minutes, and publised the corpus. We developed the procedure of automatic speech recognition, sentence extraction, segmentation/indexing, spoken retrieval and construction of lecture browsing system for classroom lecture data of our university's graduated course. These processes axe necessary to improve the usability of broadcasting sound or video data In the case of lecture, summarized and indexed lecture speech or video enables to students to more effective leaning. Our goal was to construct a framework of such structured lecture contents. To achieve this goal, first, we investigated influence of the recording methods on the speech recognition performance. It turned out that there was 23% difference on the accuracy between a high quality hand-microphone and a low quality lapel microphone. Furthermore, we improved the domain-dependent language model by using related Web texts and developed a filler insertion model. Second, we tried automatic summarization by extracting important sentences, and we obtained 0.319-0.456 κ value, comparable with human doing 0.407-0.477. Finally, we constructed the lecture browsing system which enables users to learn more effectively by using results of the procedure described above, and evaluated it
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
DOI:
--
发表时间:
2008
期刊:
IEICE Trans. Information and Systerns J91-D, 2
影响因子:
--
作者:
[S. Nakagawa, S. Togashi, M. Yamaguchi, Y. Fujii N. Kitaoka]
通讯作者:
Y. Fujii N. Kitaoka
LVCSR based on context dependent syllable acoustic models
基于上下文相关音节声学模型的 LVCSR
DOI:
--
发表时间:
2008
期刊:
影响因子:
--
作者:
[J. Zhang, L. Wang, S. Nakagawa]
通讯作者:
S. Nakagawa
講義ドキュメントのコンテンツ化と視聴システムの試作
将讲座文档转换为内容并制作查看系统原型
DOI:
--
发表时间:
2007
期刊:
第1回音声ドキュメント処理ワークショップ講演論文集
影响因子:
--
作者:
[富樫慎吾, 藤井康寿, 北岡教英, 中川聖一]
通讯作者:
中川聖一
Response timing detection using prosodic and linguistic information for human-freindly spoken dialog systems
使用韵律和语言信息进行人性化口语对话系统的响应时间检测
DOI:
--
发表时间:
2005
期刊:
人工知能学会論文誌 Vol.20, No.3
影响因子:
--
作者:
[N.Kitaoka, M.Takeuchi, R.Nishimura, S.Nakagawa]
通讯作者:
S.Nakagawa
「研究成果報告書概要(和文)」より
摘自《研究结果报告摘要(日文)》
DOI:
--
发表时间:
2005
期刊:
影响因子:
--
作者:
[Kawauchi, et. al., Nishimura et al., Dezawa et al., Yoshizawa et al., 星野 幹雄, 星野 幹雄]
通讯作者:
星野 幹雄
共 29 条
A detection method using relative phase information for spoofed speech based on speech synthesis, speaker adaptation and edited speech
-
批准号:16K12461
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.25万
-
财政年份:2016
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Study on privacy protection in spoken language
-
批准号:22650034
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.14万
-
财政年份:2010
-
负责人:NAKAGAWA Seiichi
-
依托单位:
High accuracy transcription, cleaning and fast term detection for spoken documents
-
批准号:22300059
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$11.56万
-
财政年份:2010
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A study on content summarization for large spoken documents and content retrieval through spoken dialogue
-
批准号:13480095
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$9.47万
-
财政年份:2001
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development for speech interface for form -based in formation access services on Web
-
批准号:13558033
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$4.29万
-
财政年份:2001
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Studies on Speech Recognition, Closed Caption and Summarization of Broadcast News
-
批准号:09480064
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$8.38万
-
财政年份:1997
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development of a multi-modal dialogue system and a tool for a spoken dialogue system
-
批准号:08558030
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$4.03万
-
财政年份:1996
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A study on multi-modal man-machine interface through spontaneous speech
-
批准号:06452401
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$3.39万
-
财政年份:1994
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A Research for the Formation of Basic Concepts in Physics
-
批准号:05680163
-
项目类别:Grant-in-Aid for General Scientific Research (C)
-
资助金额:$1.15万
-
财政年份:1993
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A Study on Ambiguous Utterance Understanding for Speech Input
-
批准号:03452167
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$4.54万
-
财政年份:1991
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development of a Speech Understanding system and a Spoken Dialog system
-
批准号:02555067
-
项目类别:Grant-in-Aid for Developmental Scientific Research (B)
-
资助金额:$6.78万
-
财政年份:1990
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Cooperative research on new speech recognition methods including hidden Markov models and neural networks
-
批准号:01302032
-
项目类别:Grant-in-Aid for Co-operative Research (A)
-
资助金额:$3.2万
-
财政年份:1989
-
负责人:NAKAGAWA Seiichi
-
依托单位:
海外基金