课题基金 / 基金详情

Development of a speech understanding system

Development of a speech understanding system
语音理解系统的开发
批准号:
04044108
负责人:
MIZOGUCHI Riichiro
金额:
$5.25万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for international Scientific Research
财政年份:
1992
资助国家:
日本
项目状态:
已结题
起止时间:
1992 至 1993

项目摘要

项目成果

MIZOGUCHI Riichiro的其他基金

相似基金

相关文献

中文摘要
翻译
本研究的目标是开发理解口语对话所需的基本技术,包括基于知识的语音识别系统,自然语言处理中的非单调推理和对话建模。本文的主要研究成果如下:1)验证了基于知识的韩语语音识别方法的有效性。此外,还提出了一些改进语音识别的新思想。为了避免分割困难,引入了非均匀单元。每个单元在其两端都有其静止点,中间有其瞬态部分。参数轨迹由符号表示和模糊语言变量描述。在后处理部分,对语音数据进行冗余处理,以提高识别系统的性能。原型系统对未知长度的连续韩语数字语音进行了测试,识别率达到97%。2)由于声学信息的不可靠性,连续语音的理解通常是一个坚韧。由于歧义信息的组合非常大,因此有效的搜索机制是必不可少的。在此基础上,提出了一个基于ATMS的语音理解系统框架。ATMS是一种非单调推理方法。ATMS的引入将自然语言处理的时间从64秒减少到45秒,用于理解8个日语单词的语音。3)提出了两种描述对话结构的对话模型,用于理解口语对话。一种是描述由刺激和反应组成的话语对的SR计划模型。另一种是主题分组网络(TPN),与语篇片段相对应。基于这些对话模型开发了一个预测下一个话语的机制,并在一些样本对话上进行了评估。
英文摘要
The objective of this research is development of fundamental techniques necessary to understanding spoken dialogue, which include knowledge-based speech recognition system, non-monotonic reasoning in natural language processing, and dialogue modeling. The following are the summary of the research results.1) We verified the efficiency of the knowledge-based approach for Korean speech recognition. Furthermore, some new ideas were proposed to improve the speech recognition. To avoid the difficulties in segmentation, a non-uniform unit is introduced. Every unit has its stationary point at each end of the unit, and transient part in the middle. The parameter trajectory is described by symbolic representation and fuzzy linguistic variables. Redundancy of speech data is used to improve the performance of the recognition system in the post-processor. The prototype system was tested with continuous Korean digit speech of unknown length, and the recognition rate of 97% was obtained.2) Understanding of continuous speech is generally a tough problem, since acoustic information is unreliable. An efficient search mechanism is indispensable because the combination of ambiguous information is very large. Then, we developed a framework of speech understanding system based on ATMS, which is a method of non-monotonic reasoning. The introduction of ATMS reduced elapsed time of natural language processing from 64 sec to 45 sec for understanding speech of 8 Japanese sentences.3) Two kinds of dialogue model characterizing structures in dialogue were proposed for understanding spoken dialogue. One is the SR-plan model which describes utterance pairs composed of the stimulus and the response. The other is Topic Packet Network (TPN) and corresponds to the discourse segments. A mechanism for predicting the next utterance was also developed based on these dialogue models and evaluated on some sample dialogues.
期刊论文(4)
专著(0)
科研奖励(0)
会议论文
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
山下洋一: "対話音声処理のための模擬対話の収録と分析" 日本音響学会秋季講演論文集. 23-24 (1992)
Yoichi Yamashita:“对话语音处理的模拟对话的记录和分析”日本声学学会秋季会议记录23-24(1992)。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
Yoichi Yamashita: "Next Utterance Prediction Based on Two Kinds of Dialog Models" Proceedings of Eurospeech'93. 1161-1164 (1993)
Yoichi Yamashita:“基于两种对话模型的下一个话语预测”Eurospeech93 论文集。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
Causality-compliant theory of force and motion for its innovative instruction
Development of a methodology for the next generation knowledge systems based on ontological engineering
Building a Theory-aware and Standard-compliant Knowledge Server
  • 批准号:
    19200012
  • 项目类别:
    Grant-in-Aid for Scientific Research (A)
  • 资助金额:
    $25.29万
  • 财政年份:
    2007
  • 负责人:
    MIZOGUCHI Riichiro
  • 依托单位:
Development of a Theory-Aware Authoring Workbench
  • 批准号:
    14208029
  • 项目类别:
    Grant-in-Aid for Scientific Research (A)
  • 资助金额:
    $31.7万
  • 财政年份:
    2002
  • 负责人:
    MIZOGUCHI Riichiro
  • 依托单位:
海外基金