A Study on Ambiguous Utterance Understanding for Speech Input
A Study on Ambiguous Utterance Understanding for Speech Input
批准号:
03452167
负责人:
NAKAGAWA Seiichi
金额:
$4.54万
依托单位国家:
日本
项目类别:
Grant-in-Aid for General Scientific Research (B)
财政年份:
1991
资助国家:
日本
项目状态:
已结题
起止时间:
1991 至 1993
中文摘要
本文提出了一种基于连续参数隐马尔可夫模型(HMM)的最大后验概率估计(MAPE)的无监督连续拼接训练方法。识别器采用事先建立的非特定人模型,自动生成标签序列。在连续语音识别上的实验结果表明,该模型的识别性能与监督自适应模型相当。其次,提出了一种处理感叹词和未登录词的方法,使语音识别系统能够处理对话中的自发语音。我们使用测试句子集(包括感叹词和未登录词)对语音识别系统的性能进行了评估,证实了所提方法的有效性。第三,我们研究了能够理解所有用户输入的菜单引导的口语自然语言理解系统。这项工作的动机是以下事实,即用户无法理解说什么或如何说的计算机在自然语言。系统显示一个菜单,该菜单由可接受的内容词组成,用户从菜单中选择一个词,并说出包括该词的短语。实验结果表明,该系统对于新手用户来说,性能良好。
英文摘要
We proposed an unsupervised speaker adaptation method on sequencial concatenation training that used the theory of MAPE(Maximum A Posteriori probabitity Estimation) for continuous parameter HMM.In this method, we should only specify the syllable label sequence for the utterrance. The label sequences were provided automatically by the recognizer which used a speaker-independent model in advance. The experimental results on continuous speech recognition showed that the better model gave a performance comparable to that of supervised adaptation.Secondly, we proposed a method to process interjection and unknown words so that a speech recognition system could deal with spontaneous speech in dialog. We have evaluated the peerformance of our speech recognition system using test sentence sets including interjection or unknown words, and confirmed that the proposed method worked well.Thirdly we investigated the menu-guided spoken natural language understanding system that could understand all user's inputs. This work was motivated by the following fact that a user could not understand what to say or how to say to a computer in natural language. The system displays a menu that consists of acceptable content words and the usur chooses one word from the menu and speaks out phrase that includes the word. The experimental showed that our system performed well for the novice users.
期刊论文(37)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
中川聖一: "ニューラルネットワークによる確率密度関数・事後確率の推定と母音認識" 電子情報通信学会論文誌. 76-DII. 1081-1089 (1993)
Seiichi Nakakawa:“使用神经网络估计概率密度函数和后验概率以及元音识别”电子、信息和通信工程师学会汇刊 76-DII 1081-1089 (1993)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
森屋 裕治: "対話予測を利用した音声による観光案内対話システム" 電子情報通信学会,音声研究会技術報告. SP92-121. 43-50 (1993)
Yuji Moriya:“使用对话预测的基于语音的旅游信息对话系统”,电子、信息和通信工程师研究所,语音研究组技术报告 SP92-121(1993)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
中川聖一: "ワードスポッティング法を用いた文脈自由文法制御フレーム同期型HMM連続音声認識法" 電子情報通信学会論文誌. 76-DII. 1329-1336 (1993)
Seiichi Nakakawa:“使用单词识别方法的上下文无关语法控制的帧同步 HMM 连续语音识别方法”,电子、信息和通信工程师协会学报 76-DII。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
中川聖一: "最大事後確率推定法を用いた連続出力分布型HMMの適応化" 日本音響学会誌. 49. 721-728 (1993)
Seiichi Nakakawa:“使用最大后验概率估计方法的连续输出分布类型 HMM 的自适应”日本声学学会杂志 49. 721-728 (1993)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Mikio Yamamoto: "A spoken dialog system with verification and clarification gueries" IEICE Trans.Inf.& Syst.E76-D. 84-94 (1993)
Mikio Yamamoto:“具有验证和澄清问题的语音对话系统”IEICE Trans.Inf。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 20 条
A detection method using relative phase information for spoofed speech based on speech synthesis, speaker adaptation and edited speech
-
批准号:16K12461
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.25万
-
财政年份:2016
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Study on privacy protection in spoken language
-
批准号:22650034
-
项目类别:Grant-in-Aid for Challenging Exploratory Research
-
资助金额:$2.14万
-
财政年份:2010
-
负责人:NAKAGAWA Seiichi
-
依托单位:
High accuracy transcription, cleaning and fast term detection for spoken documents
-
批准号:22300059
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$11.56万
-
财政年份:2010
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Automatic indexing for lecture speech and its advanced utilization through speech interaction
-
批准号:17300064
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$10.26万
-
财政年份:2005
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A study on content summarization for large spoken documents and content retrieval through spoken dialogue
-
批准号:13480095
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$9.47万
-
财政年份:2001
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development for speech interface for form -based in formation access services on Web
-
批准号:13558033
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$4.29万
-
财政年份:2001
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Studies on Speech Recognition, Closed Caption and Summarization of Broadcast News
-
批准号:09480064
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$8.38万
-
财政年份:1997
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development of a multi-modal dialogue system and a tool for a spoken dialogue system
-
批准号:08558030
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$4.03万
-
财政年份:1996
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A study on multi-modal man-machine interface through spontaneous speech
-
批准号:06452401
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$3.39万
-
财政年份:1994
-
负责人:NAKAGAWA Seiichi
-
依托单位:
A Research for the Formation of Basic Concepts in Physics
-
批准号:05680163
-
项目类别:Grant-in-Aid for General Scientific Research (C)
-
资助金额:$1.15万
-
财政年份:1993
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Development of a Speech Understanding system and a Spoken Dialog system
-
批准号:02555067
-
项目类别:Grant-in-Aid for Developmental Scientific Research (B)
-
资助金额:$6.78万
-
财政年份:1990
-
负责人:NAKAGAWA Seiichi
-
依托单位:
Cooperative research on new speech recognition methods including hidden Markov models and neural networks
-
批准号:01302032
-
项目类别:Grant-in-Aid for Co-operative Research (A)
-
资助金额:$3.2万
-
财政年份:1989
-
负责人:NAKAGAWA Seiichi
-
依托单位:
海外基金