Studies on Multimodal Communication by Integrating Speech and Diagram
Studies on Multimodal Communication by Integrating Speech and Diagram
批准号:
08458078
负责人:
DOSHITA Shuji
金额:
$4.67万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
1996
资助国家:
日本
项目状态:
已结题
起止时间:
1996 至 1997
中文摘要
本研究的目的是通过实现语音输入的多模态绘图系统,实现人机之间的多模态交流。本研究的主要结果如下:1.语音、手势和手势的集成我们开发了多模式集成方法,而无需假设每个输入元素的同步。通过评价多模态听写系统中信息输入准确性的提高,检验了该集成方法的有效性.利用语音和指向手势实现多模态接口我们开发了协作机制,该机制使用指向手势识别的结果,这是可靠的通道,用于识别语音,这是相对不可靠的通道。利用手势识别的结果,语音识别的准确率提高了6%左右。利用这种集成方法,我们实现了一个语音输入的多通道绘图系统.多模态绘图系统中基于图形和上下文信息的话语理解我们通过堆栈实现了使用上下文管理机制的话语理解机制,以处理自发语音(例如:G.省略号、引用表达式等)。我们将这种话语理解机制集成到多模态绘图系统中。
英文摘要
The aim of this research is to realize multimodal communication between human and machines through implementing multimodal drawing system with speech input. The results of this research are below :1. Integration of speech, gesture and diagramWe developed multi-mode integration method without assuming the synchronization of each input elements. The validity of this integration method was examined by evaluating the improvement of accuracy of information input in multimodal dictation system.2. Implementation of multimodal interface using speech and pointing gestureWe developed the cooperation mechanism which uses the result of pointing gesture recognition, which is reliable channel, for the recognition of speech, which is relatively unreliable channel. By using the result of pointing gesture recognition, the accuracy of speech recognition raised about 6%. Using this integration method, we implemented a multimodal drawing system with speech input.3. Interpretation of utterance using diagrammatic and contextual information in multimodal drawing system We realized utterance understanding mechanism which uses context management mechanism by stack in order to deal with spontaneous speech (e. g. ellipsis, reference expression, etc.). We integrated this utterance understanding mechanism to multimodal drawing system.
期刊论文(28)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Doshita,S.: "UNDERSTANDING AND GENERATING DIALOGUE BY INTEGRATING PRO-CESSING OF SPEECH,LANGUAE AND CONCEPT" Proc.of International Symposium on Spoken Dialogue '96. 1-8 (1996)
Doshita,S.:“通过整合语音、语言和概念处理来理解和生成对话”Proc.of 国际口语研讨会 96。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Araki, M.: "Automatic Evaluation Environment for Spoken Dialogue Systems" in Mayr, E.et al.ed. "Dialogue Processing in Spoken Language Systems". (1997)
Araki, M.:“口语对话系统的自动评估环境”,Mayr, E.et al.ed。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Doshita, S.: "Human-machine communication by speech" Ohmsha inc.(1998)
Doshita, S.:“通过语音进行人机通信”Ohmsha inc.(1998)
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
荒木雅弘: "図像情報を利用した講演調音声のディクテーション" 第53回情報処理学会全国大会講演論文集. 357-358 (1996)
Masahiro Araki:“使用图像信息听写类似讲座的演讲”第 53 届日本信息处理学会全国会议记录 357-358(1996 年)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Kawahara, T.: "Speaking-style dependent lexicalized filler model for key-phrase detection and verification" 電子情報通信学会技術研究報告. SP97-78. (1997)
Kawahara, T.:“用于关键短语检测和验证的说话风格相关词汇填充模型”SP97-78。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 23 条
Research on Understanding and Generating Dialogue by Integrated Processing of Speech, Language and Concept
-
批准号:05241103
-
项目类别:Grant-in-Aid for Scientific Research on Priority Areas
-
资助金额:$16.64万
-
财政年份:1996
-
负责人:DOSHITA Shuji
-
依托单位:
Robust speech understanding against inter-speaker variation and ungrammatical utterances based on high-accuracy speech recognition and semantic driven parsing method
-
批准号:05452357
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$3.33万
-
财政年份:1993
-
负责人:DOSHITA Shuji
-
依托单位:
Intelligent pattern recognition and understanding by integrating probabilistic and symbolic reasoning
-
批准号:02452281
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$4.1万
-
财政年份:1990
-
负责人:DOSHITA Shuji
-
依托单位:
Fundamental Research of Speech Translation Based on High Accurate Speech Recognition and Language-Concept Understanding
-
批准号:62420052
-
项目类别:Grant-in-Aid for General Scientific Research (A)
-
资助金额:$14.02万
-
财政年份:1987
-
负责人:DOSHITA Shuji
-
依托单位:
海外基金