Interactive task learning via embodied corrective feedback
Interactive task learning via embodied corrective feedback
复制标题
通过具体的纠正反馈进行交互式任务学习
DOI:
--
复制
发表时间:
2020
影响因子:
1.9
通讯作者:
A. Lascarides
中科院分区:
文献类型:
--
作者:
Mattias Appelgren;A. Lascarides
This paper addresses a task in Interactive Task Learning (Laird et al. IEEE Intell Syst 32:6–21, 2017). The agent must learn to build towers which are constrained by rules, and whenever the agent performs an action which violates a rule the teacher provides verbal corrective feedback: e.g. “No, red blocks should be on blue blocks”. The agent must learn to build rule compliant towers from these corrections and the context in which they were given. The agent is not only ignorant of the rules at the start of the learning process, but it also has a deficient domain model, which lacks the concepts in which the rules are expressed. Therefore an agent that takes advantage of the linguistic evidence must learn the denotations of neologisms and adapt its conceptualisation of the planning domain to incorporate those denotations. We show that by incorporating constraints on interpretation that are imposed by discourse coherence into the models for learning (Hobbs in On the coherence and structure of discourse, Stanford University, Stanford, 1985; Asher et al. in Logics of conversation, Cambridge University Press, Cambridge, 2003), an agent which utilizes linguistic evidence outperforms a strong baseline which does not.
DOI:
--
发表时间:
2019
期刊:
--
影响因子:
--
作者:
Appelgren M
通讯作者:
Appelgren M
DOI:
--
发表时间:
2018
期刊:
--
影响因子:
--
作者:
Hristov YS
通讯作者:
Hristov YS
DOI:
--
发表时间:
2019
期刊:
--
影响因子:
--
作者:
Appelgren M
通讯作者:
Appelgren M