Automatic Anaphora Resolution for Norwegian (ARN)

Automatic Anaphora Resolution for Norwegian (ARN)
复制标题

挪威语自动照应解析 (ARN)

DOI:
10.1007/978-3-540-71412-5_11
复制
发表时间:
2007
期刊:
--
影响因子:
--
通讯作者:
Gordana Ilic Holen
Gordana Ilic Holen
中科院分区:
--
文献类型:
--
作者:
Gordana Ilic Holen

文献摘要

被引文献

相似文献

Thearnsystem——挪威语自动照应解析系统——是一个基于规则的照应解析系统,它是在英语语言的两个现有系统的基础上设计的:Mitkov 的原始方法及其后来的发展,以及 Lappin 和 Leass 的治疗系统。这些系统中的大量规则基于中心理论支持的超级规则,该规则在状语和介词短语中优先考虑候选主语而不是候选宾语,并且优先考虑候选宾语。由于挪威语和英语之间的信息结构存在差异,这些规则不适用于挪威语。尽管两种语言都倾向于避免在该主题上传达新信息,但挪威语却竭尽全力避免这种情况。这种趋势导致挪威语中带有咒骂主语的句子数量比英语中要多得多,使得这些主语不适合作为先行词候选。做出复杂的偏好来处理性别/性别冲突并优先考虑代词候选者和与照应词接近的候选者已被证明是挪威语的一个很好的策略。arn旨在解决除代词'it(中性)'之外的第三人称代词,并且已经达到了70.5%。
Thearnsystem — an Automatic Anaphora Resolution System for Norwegian — is a rule-based anaphora resolution system that was designed on the basis of two existing systems for the English language: Mitkov’s Original Approach with its later developmentmars, and therapsystem by Lappin and Leass. A substantial group of rules within these systems is based upon a super-rule supported by Centering theory, which gives preference to subjects candidates over objects candidates, and object candidates over candidates within adverbial and prepositional phrases. These rules cannot be applied to Norwegian, due to differences in information structure between Norwegian and English. Although there is a tendency in both languages to avoid conveying new information with the subject, Norwegian goes to much greater lengths to avoid it. This tendency leads to a substantially higher number of sentences with expletive subjects in Norwegian than in English, rendering those subjects unsuitable as antecedent candidates.Making a complex preference to handle the sex/gender conflict and giving preference to pronominal candidates and candidates in close proximity to the anaphor has proved to be a good strategy for Norwegian.arnwas designed to resolve the third person pronoun with the exception of pronoundet’it (neut.)’, and has achieved an accuracy of 70.5%.