An Annotated Corpus of Reference Resolution for Interpreting Common Grounding

An Annotated Corpus of Reference Resolution for Interpreting Common Grounding
复制标题

DOI:
10.1609/aaai.v34i05.6442
复制
发表时间:
2019-11
期刊:
ArXiv
影响因子:
--
通讯作者:
Takuma Udagawa;Akiko Aizawa
Takuma Udagawa;Akiko Aizawa
中科院分区:
其他
文献类型:
--
作者:
Takuma Udagawa;Akiko Aizawa

文献摘要

相似文献

共同基础是创建、修复和更新相互理解的过程,这是自然语言对话的一个基本方面。然而,解释共同基础的过程是一项具有挑战性的任务,特别是在连续和部分可观察的背景下,复杂的模糊性,不确定性,部分理解和误解。当我们处理仍然具有有限的自然语言理解和生成能力的对话系统时,解释变得更具挑战性。为了解决这个问题,我们认为参考决议的共同基础的中心子任务,并提出了一个新的资源来研究其中间过程。基于一个简单而通用的注释模式,我们从现有语料库中收集了5,191个对话中的40,172个指称表达,沿着对指称解释的多种判断。我们表明,我们的注释是高度可靠的,通过合理的分歧的自然程度捕捉到共同基础的复杂性,并允许更详细和定量的分析共同基础的战略。最后,我们展示了我们的注释的优势,解释,分析和改善基线对话系统的共同基础。
Common grounding is the process of creating, repairing and updating mutual understandings, which is a fundamental aspect of natural language conversation. However, interpreting the process of common grounding is a challenging task, especially under continuous and partially-observable context where complex ambiguity, uncertainty, partial understandings and misunderstandings are introduced. Interpretation becomes even more challenging when we deal with dialogue systems which still have limited capability of natural language understanding and generation. To address this problem, we consider reference resolution as the central subtask of common grounding and propose a new resource to study its intermediate process. Based on a simple and general annotation schema, we collected a total of 40,172 referring expressions in 5,191 dialogues curated from an existing corpus, along with multiple judgements of referent interpretations. We show that our annotation is highly reliable, captures the complexity of common grounding through a natural degree of reasonable disagreements, and allows for more detailed and quantitative analyses of common grounding strategies. Finally, we demonstrate the advantages of our annotation for interpreting, analyzing and improving common grounding in baseline dialogue systems.