Asynchronous Multimodal Text Entry Using Speech and Gesture Keyboards

Asynchronous Multimodal Text Entry Using Speech and Gesture Keyboards
复制标题

使用语音和手势键盘的异步多模式文本输入

DOI:
--
复制
发表时间:
2011
期刊:
Interspeech
影响因子:
--
通讯作者:
K. Vertanen
K. Vertanen
中科院分区:
--
文献类型:
--
作者:
P. Kristensson;K. Vertanen

文献摘要

被引文献

相似文献

我们建议通过结合语音和手势键盘输入来减少文本输入中的错误。我们描述了一种以异步和灵活的方式将识别结果组合在一起的合并模型。我们收集了输入简短电子邮件句子和网络搜索查询的用户的语音和手势数据。通过合并两种模式的识别结果,电子邮件句子的单词错误率相对降低了53%,网络搜索的单词错误率相对降低了29%。对于有语音错误的电子邮件话语,我们研究了只提供错误单词的手势键盘更正。在没有用户明确指出错误单词的情况下,我们的模型能够将单词错误率相对降低44%。索引术语:手机文本输入、多模式界面
We propose reducing errors in text entry by combining speech and gesture keyboard input. We describe a merge model that combines recognition results in an asynchronous and flexible manner. We collected speech and gesture data of users entering both short email sentences and web search queries. By merging recognition results from both modalities, word error rate was reduced by 53% relative for email sentences and 29% relative for web searches. For email utterances with speech errors, we investigated providing gesture keyboard corrections of only the erroneous words. Without the user explicitly indicating the incorrect words, our model was able to reduce the word error rate by 44% relative. Index Terms: mobile text entry, multimodal interfaces