Annotating Reference and Coreference In Dialogue Using Conversational Agents in games
Annotating Reference and Coreference In Dialogue Using Conversational Agents in games
批准号:
EP/W001632/1
负责人:
Massimo Poesio
金额:
$139.06万
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2022
资助国家:
英国
项目状态:
未结题
起止时间:
2022 至 --
中文摘要
点击翻译按钮获取中文摘要
英文摘要
The development of modern neural network architectures architectures such as the encoder/decoder model and the Transformer has brought about an explosion of interest in neural models for AI systems able to engage in conversations (aka conversational agents), reflected by a spike of published work, dedicated workshops, and industry-sponsored competitions and grants. While at first these models were applied to simple chatbots, the focus of research has been shifting towards conversational agents capable of engaging in more complex and task-oriented dialogue such as restaurant booking or question answering. But the results on these tasks show that while end-to-end architectures without dedicated models for semantic interpretation can work well for chatbots, conversational agents carrying out more complex tasks require greater ablity to handle such aspects of interpretation, and some form of modelling of context. Among the aspects of natural language interpretation that require more advanced architectures are COREFERENCE and REFERENCE. For an example of the importance of coreference in dialog, consider the following except from a real-life chat conversation, where both participants continually use anaphoric expressions such as BOTH, THEY, IT, etc to refer to previously introduced entities such as Google or Microsoft.A:Are you a fan of Google or Microsoft?B:Both are excellent technology they are helpful in many ways. For the security purpose both are super.A:I'm not a huge fan of Google, but I use it a lot because I have to. I think they are a monopoly in some sense.B:Google provides online related services and products, which includes search engine and cloud computing.A:Yeah, their services are good. I'm just not a fan of intrusive they can be on our personal livesEnriching conversational agents with the ability to carry out these forms of interpretation raises two issues. First, developing models for these tasks requires specific training data: most deep-learning architectures are trained on large amounts of freely available written text. Training a coreference resolver on written text and domain-adapting it to dialogue however has proven ineffective as coreference in dialogue involves different phenomena and is more involved than coreference in text. Second, the developed architectures require specific modules that enable them to interpret coreference and reference. Our group has pioneered the use of Games-With-A-Purpose (GWAPs) to collect data for NLP, resulting in the largest NLP dataset collected using GWAPs or indeed crowdsourcing. But there is a fundamental difference between conversation and written text: the latter is designed to be read by third parties, whereas research has shown that overhearers to a conversation only acquire a partial understanding of what was said.OUR PROPOSED SOLUTION to the problem of creating large annotated datasets of coreference and reference interpretation in conversation is to collect the judgments for anaphoric and referential information via GAMES IN WHICH CONVERSATIONAL AGENTS INTERACT WITH HUMAN PLAYERS AND EVOLVE BY ACQUIRING INFORMATION FROM THEM. This idea builds on recent work by Facebook and Microsoft, among others, that pioneered the use of conversational agents in games to collect data about dialogue, and of Hockenmaier and her lab. Our agents will be deployed in gaming platforms such as LIGHT and MINECRAFT in collaboration with these labs. But whereas in previous work conversational agents only interact with the aim to improve their end-to-end behavior, in the proposed project we will develop artificial agents able to improve their ability to interpret coreference and reference by collecting judgments about these interpretation aspects via CLARIFICATION QUESTIONS to the players at appropriate moments, which can also be used to annotate a dataset.
期刊论文(10)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
The CODI-CRAC 2022 Shared Task on Anaphora, Bridging, and Discourse Deixis in Dialogue
CODI-CRAC 2022 对话中的照应、桥接和话语指示语共享任务
DOI:
--
发表时间:
2022
期刊:
影响因子:
--
作者:
[Yu, J]
通讯作者:
Yu, J
Aggregating crowdsourced and automatic judgments to scale up a corpus of anaphoric reference for fiction and Wikipedia texts
聚合众包和自动判断,以扩大小说和维基百科文本的照应参考语料库
DOI:
--
发表时间:
2023
期刊:
影响因子:
--
作者:
[Yu, J]
通讯作者:
Yu, J
LingoTowns: A Virtual World For Natural Language Annotation and Language Learning
LingoTowns:自然语言注释和语言学习的虚拟世界
DOI:
10.1145/3505270.3558323
发表时间:
2022
期刊:
影响因子:
--
作者:
[Madge C]
通讯作者:
Madge C
Coreference Annotation of an Arabic Corpus using a Virtual World Game
使用虚拟世界游戏对阿拉伯语语料库进行共指注释
DOI:
10.18653/v1/2022.wanlp-1.37
发表时间:
2022
期刊:
影响因子:
--
作者:
[Aliady W]
通讯作者:
Aliady W
ARCIDUCA: Annotating Reference and Coreference In Dialogue Using Conversational Agents in games
ARCIDUCA:在游戏中使用会话代理注释对话中的参考和共指
DOI:
--
发表时间:
2022
期刊:
影响因子:
--
作者:
[Poesio, M]
通讯作者:
Poesio, M
共 8 条
Creating anaphorically annotated resources through semantic wikis (AnaWiki)
-
批准号:EP/F00575X/1
-
项目类别:Research Grant
-
资助金额:$18.26万
-
财政年份:2007
-
负责人:Massimo Poesio
-
依托单位:
海外基金