Overview of the Patent Translation Task at the NTCIR-7 Workshop
Overview of the Patent Translation Task at the NTCIR-7 Workshop
复制标题
DOI:
--
复制
发表时间:
2008
期刊:
影响因子:
--
通讯作者:
Atsushi Fujii;M. Utiyama;Mikio Yamamoto;T. Utsuro
中科院分区:
文献类型:
--
作者:
Atsushi Fujii;M. Utiyama;Mikio Yamamoto;T. Utsuro
To aid research and development in machine translation, we have produced a test collection for Japanese/English machine translation and performed the Patent Translation Task at the Seventh NTCIR Workshop. To obtain a parallel corpus, we extracted patent documents for the same or related inventions published in Japan and the United States. Our test collection includes approximately 2 000 000 sentence pairs in Japanese and English, which were extracted automatically from our parallel corpus. These sentence pairs can be used to train and evaluate machine translation systems. Our test collection also includes search topics for cross-lingual patent retrieval, which can be used to evaluate the contribution of machine translation to retrieving patent documents across languages. This paper describes our test collection, methods for evaluating machine translation, and evaluation results for research groups participated in our task. Our research is the first significant exploration into utilizing patent information for the evaluation of machine translations.