Clause recognition in the framework of alignment

Clause recognition in the framework of alignment
复制标题

对齐框架中的子句识别

DOI:
10.1075/cilt.136.35pap
复制
发表时间:
1997
期刊:
影响因子:
--
通讯作者:
Haris Papageorgiou
Haris Papageorgiou
中科院分区:
--
文献类型:
--
作者:
Haris Papageorgiou

文献摘要

被引文献

相似文献

本文探讨了使用词性信息(作为基于统一规则的词性标记器的输出)、CMTAG模块(主要用于修正早期处理中的错误)和基于语言规则的句法分析器来实现无限制文本的可靠从句识别的可能性。简单/复杂从句的识别被认为是平行文本双语对齐的一个基本组成部分。这项工作的重点之一是处理超长句子的能力。句法分析器能够对子句结构进行分析和标注。该系统在实验语料库上得到了应用。我们所取得的结果是非常有希望的。
In this paper we explore the possibility of achieving reliable clause identification of unrestricted text by using POS information (as the output of a unification rule-based part of speech tagger) a CMTAG module trying mainly to fix errors from earlier processing and a linguistic rule-based parser. Identification of simple/complex clauses is considered here as a basic component in the framework of bilingual alignment of parallel texts. One of the important points of this work is the ability for processing very long sentences. Parser is capable of analysing and labeling clause structure. The system is applied to an experimental corpus. The results we have obtained are very promising.