Collecting machine-translation-aided bilingual dialogues for corpus-based speech translation

Collecting machine-translation-aided bilingual dialogues for corpus-based speech translation
复制标题

收集机器翻译辅助的双语对话以进行基于语料库的语音翻译

DOI:
10.21437/eurospeech.2003-735
复制
发表时间:
2003
期刊:
2007 IEEE International Conference on Acoustics, Speech and Signal Processing - ICASSP '07
影响因子:
--
通讯作者:
G. Kikui
G. Kikui
中科院分区:
--
文献类型:
--
作者:
T. Takezawa;G. Kikui

文献摘要

被引文献

相似文献

ATR口语翻译研究实验室正在建设一个庞大的英语和日语双语语料库,以提高语音翻译技术,以便人们可以在出国旅游、餐饮和购物以及酒店情况下使用便携式翻译系统。作为语料库建设活动的一部分,我们一直在使用一个实验性的英语和日语之间的翻译系统来收集对话数据。数据收集的目的是研究在这类系统面前的交际行为和语言表达方式。我们使用人工打字员转录用户的话语,并将其输入到英语和日语之间的机器翻译系统,而不是使用语音识别系统。在本文中,我们概述了我们的活动,并根据基本特征进行了讨论。
A huge bilingual corpus of English and Japanese is being built at ATR Spoken Language Translation Research Laboratories in order to enhance speech translation technology, so that people can use a portable translation system for traveling abroad, dining and shopping, as well as hotel situations. As a part of these corpus construction activities, we have been collecting dialogue data using an experimental translation system between English and Japanese. The purpose of this data collection is to study the communication behaviors and linguistic expressions preferred in front of such systems. We use human typists to transcribe the users’ utterances and input them into a machine translation system between English and Japanese instead of using speech recognition systems. In this paper, we present an overview of our activities and discussionsbased on the basic characteristics.