Exploiting Speech Understanding in Intelligent Interfaces
Exploiting Speech Understanding in Intelligent Interfaces
批准号:
06044055
负责人:
WARD Nigel
金额:
$2.69万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for international Scientific Research
财政年份:
1994
资助国家:
日本
项目状态:
已结题
起止时间:
1994 至 1995
中文摘要
点击翻译按钮获取中文摘要
英文摘要
We are interested in the use of spoken language in human-computer interaction. The inspiration is the fact that, for human-human interaction, meaningful exchanges can take place even without accurate recognition of the words the other is saying --- this being possible due to shared knowledge and complementary communication channels, especially gesture and prosody. We want to exploit this fact for man-machine interfaces.Therefore we are doing three things :1. Using simple speech recognition to augment graphical user interfaces, well integrated with other input modalities : keyboard, mouse, and touch screen.2. Building systems able to engage in simple conversations, using mostly prosodic clues. To sketch out our latest success :We conjectured that it would be possible for Japanese to decide when to produce many back-channel utterances based on prosodic clues alone, without reference to meaning.We found thatneither vowel lengthening, volume changes, nor energy level (to detect when the other finished speaking) were by themselves good predictors of when to produce an aizuchi. The best predictor was a low pitch level.Specifically, upon detection of the end of a region of pitch less than.9 times the local median pitch and continuing for 150ms, coming after at least 600ms of speech, the system predicted an aizuchi 200ms to 300ms later, providing it had not done so within the preceding 1 second.We also built a real-time system based on the above decision rule. A human stooge steered the conversation to a suitable topic and then switched on the system. After swich-on the stooge's utterances and the system's outputs, mixed together, produced one side of the conversation. We found that none of the 5 subjects had realized that his conversation partner had become partially automated.3. Building tools and collecting data to help do 1 and 2.
期刊论文(19)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Tajchman,Gary and Dan,Jurafsky and Eric Folder: "Learning Phonological Rule Probabilities from Speech Corpora with Exploratory Computational Phonology" In Proceedings of ACL95. 9-15 (1995)
Tajchman、Gary 和 Dan、Jurafsky 和 Eric Folder:“通过探索性计算音系学从语音语料库学习音系规则概率”,ACL95 论文集。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Gildea, Daniel and Daniel.Jurafsky: "Learning Bias and Phonological Rules Induction" Computational Linguistics. (1995)
Gildea、Daniel 和 Daniel.Jurafsky:“学习偏差和语音规则归纳”计算语言学。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Nigel, WARD: "Using Prosodic Clucs to Decide When to Produce Back-Channel Utterances" CSLP.
Nigel,WARD:“使用韵律线索来决定何时产生 Back-Channel 话语”CSLP。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Jurafsky, Daniel: "A Probabilistic Model of Lexical and Syntactic Access and Disambiguation" Cognitive Science.
Jurafsky,丹尼尔:“词汇和句法访问和消歧的概率模型”认知科学。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
Nigel Ward: "An Approach to Tightly-Coupled Syntactic/Semantic Processing for Speech Understanding" Proceedings of the AAAT Workshop on the Integration of Natural Language and Speech Processing. 50-57 (1994)
Nigel Ward:“用于语音理解的紧耦合句法/语义处理方法”自然语言与语音处理集成 AAAT 研讨会论文集。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 15 条
外国語会話能力養成のための対話的反射訓練システム
-
批准号:12040209
-
项目类别:Grant-in-Aid for Scientific Research on Priority Areas (A)
-
资助金额:$1.34万
-
财政年份:2000
-
负责人:WARD Nigel
-
依托单位:
Non-lexical Sounds : a New Interface Modality for Voice-based Information Delivery Systems
-
批准号:11680412
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.6万
-
财政年份:1999
-
负责人:WARD Nigel
-
依托单位:
実時間音声理解応答を利用した機械操作における指動作訓練支援システムの研究
-
批准号:08750301
-
项目类别:Grant-in-Aid for Encouragement of Young Scientists (A)
-
资助金额:$0.7万
-
财政年份:1996
-
负责人:WARD Nigel
-
依托单位:
海外基金