Hypothesis Selection in Machine Transliteration: A Web Mining Approach
Hypothesis Selection in Machine Transliteration: A Web Mining Approach
复制标题
机器音译中的假设选择:一种网络挖掘方法
DOI:
--
复制
发表时间:
2008
期刊:
影响因子:
--
通讯作者:
H. Isahara
中科院分区:
文献类型:
--
作者:
Jong;H. Isahara
We propose a new method of selecting hypotheses for machine transliteration. We generate a set of Chinese, Japanese, and Korean transliteration hypotheses for a given English word. We then use the set of transliteration hypotheses as a guide to finding relevant Web pages and mining contextual information for the transliteration hypotheses from the Web page. Finally, we use the mined information for machine-learning algorithms including support vector machines and maximum entropy model designed to select the correct transliteration hypothesis. In our experiments, our proposed method based on Web mining consistently outperformed systems based on simple Web counts used in previous work, regardless of the language.