A study on content summarization for large spoken documents and content retrieval through spoken dialogue
A study on content summarization for large spoken documents and content retrieval through spoken dialogue
批准号:
13480095
负责人:
NAKAGAWA Seiichi
金额:
$9.47万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2001
资助国家:
日本
项目状态:
已结题
起止时间:
2001 至 2004
中文摘要
为了开发一个面向开放领域口语文档检索的准确的大词汇量连续语音识别系统,我们提出了一种并行使用两种搜索算法的搜索方法,以实现高效和准确的解码。我们对这种新的搜索算法进行了评估,在不增加计算代价的情况下,识别性能得到了显着的提高。我们还提出了将机器学习技术应用于组合多个LVCSR模型的输出的任务。与罗孚等投票方案相比,所提出的技术具有优势,特别是在大多数参与模型不可靠的情况下。通过使用该技术,我们完成了语音驱动的Web检索任务,提高了语音查询的语音识别准确率,从而提高了语音驱动的Web检索的准确率。为此,我们研究了语言表层信息与人类结果之间的关系,得到了有用的表层语言信息。接下来,我们根据这些信息对口语讲座进行了总结,并将其与人类的结果进行了比较。结果,我们得到了与人的结果相当的更好的F-测度和k值。我们在一个口语对话系统中开发了一个可移植的语音识别模块和一个翻译模块。此外,我们还开发了一个对话策略设计工具,并将其应用于富士山观光指南检索、文献检索和酒店预订检索,验证了该工具的实用性。
英文摘要
To develop an accurate large vocabulary continuous speech recognition system for spoken document retrieval in open domain, we proposed a search method using two search algorithms in parallel to achieve efficient and accurate decoding. We evaluated this new search algorithm and obtained significant improvement of recognition performance without severe increase of computational cost We also proposed to apply machine learning techniques to the task of combining outputs of multiple LVCSR models. The proposed technique had advantages over that by voting schemes such as ROVER, especially when the majority of participating models are not reliable. By using this technique, we performed a speech-driven Web retrieval task and improved speech recognition accuracy of spoken queries and then improved retrieval accuracy in speech driven Web retrieval We tried the summarization of spoken lectures. For this purpose, we investigated relations between linguistic surface information and human's results, and we obtained useful surface linguistic information. Next, we summarized spoken lectures based on this information, and compared them with human's results. As a result, we obtained a better F-measure and k-value comparable with human's results. We have developed a portable speech recognition module and an interpreter module in a spoken dialogue system. Furthermore, we also developed a dialogue strategy design tool, applied it to Mt.Fuji sightseeing guidance retrieval, literature retrieval and hotel reservation retrieval and then confirmed the usefulness.
期刊论文(56)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
音声認識誤りと未知語に頑健な音声文書検索手法
针对语音识别错误和未知词的鲁棒性语音文档检索方法
DOI:
--
发表时间:
2003
期刊:
電子情報通信学会論文誌 86-DII・10
影响因子:
--
作者:
[西崎博光]
通讯作者:
西崎博光
Satoshi Kobayashi: "Extracting summarizing of lectures based on linguistic surface and prosodic information"Proc.Workshop on Spontaneous Speech Processing and recognition. 211-214 (2003)
Satoshi Kobayashi:“基于语言表面和韵律信息提取讲座摘要”Proc. 自发语音处理和识别研讨会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
機械学習を用いた複数の大語彙連続音声認識モデルの出力の混合
使用机器学习混合多个大词汇量连续语音识别模型的输出
DOI:
--
发表时间:
2004
期刊:
電子情報通信学会論文誌 87-DII・7
影响因子:
--
作者:
[C.Nattee, 宇津呂武仁]
通讯作者:
宇津呂武仁
Detection and recognition of correction Utterances on miss-recognition of spoken dialog system.
口语对话系统误识别的纠正话语检测与识别。
DOI:
--
发表时间:
2004
期刊:
Trans.Inst.Elect.Comm.Inform. 87-D II・7
影响因子:
--
作者:
[Norihide Kitaoka]
通讯作者:
Norihide Kitaoka
Masamitsu Umeda: "Interpreter for highly portable spoken dialogue system"Proc.4-th Sigdial Workshop on discourse and Dialogue. 105-114 (2003)
Masamitsu Umeda:“高度便携口语对话系统的翻译”Proc.4-th Sigdial 话语和对话研讨会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 27 条
A detection method using relative phase information for spoofed speech based on speech synthesis, speaker adaptation and edited speech
-
批准号:16K12461
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.25万
-
财政年份:2016
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Study on privacy protection in spoken language
-
批准号:22650034
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.14万
-
财政年份:2010
-
负责人:NAKAGAWA Seiichi
-
依托单位:
High accuracy transcription, cleaning and fast term detection for spoken documents
-
批准号:22300059
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$11.56万
-
财政年份:2010
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Automatic indexing for lecture speech and its advanced utilization through speech interaction
-
批准号:17300064
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$10.26万
-
财政年份:2005
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development for speech interface for form -based in formation access services on Web
-
批准号:13558033
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$4.29万
-
财政年份:2001
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Studies on Speech Recognition, Closed Caption and Summarization of Broadcast News
-
批准号:09480064
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$8.38万
-
财政年份:1997
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development of a multi-modal dialogue system and a tool for a spoken dialogue system
-
批准号:08558030
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$4.03万
-
财政年份:1996
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A study on multi-modal man-machine interface through spontaneous speech
-
批准号:06452401
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$3.39万
-
财政年份:1994
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A Research for the Formation of Basic Concepts in Physics
-
批准号:05680163
-
项目类别:Grant-in-Aid for General Scientific Research (C)
-
资助金额:$1.15万
-
财政年份:1993
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A Study on Ambiguous Utterance Understanding for Speech Input
-
批准号:03452167
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$4.54万
-
财政年份:1991
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development of a Speech Understanding system and a Spoken Dialog system
-
批准号:02555067
-
项目类别:Grant-in-Aid for Developmental Scientific Research (B)
-
资助金额:$6.78万
-
财政年份:1990
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Cooperative research on new speech recognition methods including hidden Markov models and neural networks
-
批准号:01302032
-
项目类别:Grant-in-Aid for Co-operative Research (A)
-
资助金额:$3.2万
-
财政年份:1989
-
负责人:NAKAGAWA Seiichi
-
依托单位:
海外基金