Model-based and model-free Pavlovian reward learning: revaluation, revision, and revelation.

Model-based and model-free Pavlovian reward learning: revaluation, revision, and revelation.
复制标题

基于模型和无模型的Pavlovian奖励学习:重估,修订和启示。

DOI:
10.3758/s13415-014-0277-8
复制
发表时间:
2014-06
期刊:
Cognitive, affective & behavioral neuroscience
影响因子:
--
通讯作者:
Berridge KC
Berridge KC
中科院分区:
其他
文献类型:
--
作者:
Dayan P;Berridge KC

文献摘要

参考文献

被引文献

相似文献

证据支持至少两种方法来学习奖励和惩罚并做出预测以指导行动。一种称为无模型的方法从回顾性经验中逐步获取对环境和行动的长期价值的缓存估计。另一种方法称为基于模型的方法,它使用对环境、期望和预期计算的表征来对未来价值进行认知预测。这两种方法在工具学习的计算分析中得到了广泛的关注。相比之下,虽然缺乏完整的计算分析,但巴甫洛夫学习和预测通常被认为是完全无模型的。在这里,我们修改了这一假设,并回顾了令人信服的证据,从巴甫洛夫重估实验表明,巴甫洛夫预测可以涉及自己的形式的基于模型的评估。在基于模型的巴甫洛夫评估中,身体和大脑的主要状态影响价值计算,从而产生强大的激励动机,有时可能是相当新的。我们认为,这种修订后的巴甫洛夫观点的预测,响应和选择的计算景观的后果。我们还重新审视了巴甫洛夫和工具性学习在控制激励动机方面的差异。
Evidence supports at least two methods for learning about reward and punishment and making predictions for guiding actions. One method, called model-free, progressively acquires cached estimates of the long-run values of circumstances and actions from retrospective experience. The other method, called model-based, uses representations of the environment, expectations and prospective calculations to make cognitive predictions of future value. Extensive attention has been paid to both methods in computational analyses of instrumental learning. By contrast, although a full computational analysis has been lacking, Pavlovian learning and prediction has typically been presumed to be solely model-free. Here, we revise that presumption and review compelling evidence from Pavlovian revaluation experiments showing that Pavlovian predictions can involve their own form of model-based evaluation. In model-based Pavlovian evaluation, prevailing states of the body and brain influence value computations, and thereby produce powerful incentive motivations that can sometimes be quite new. We consider the consequences of this revised Pavlovian view for the computational landscape of prediction, response and choice. We also revisit differences between Pavlovian and instrumental learning in the control of incentive motivation.
DOI: 10.1523/jneurosci.0897-08.2008
发表时间: 2008-05-28
影响因子: 5.3
作者:
Bray, Signe;Rangel, Antonio;O'Doherty, John P.
通讯作者: O'Doherty, John P.
DOI: 10.1016/j.neuron.2009.05.014
发表时间: 2009-06-11
期刊: NEURON
影响因子: 16.2
作者:
Boorman, Erie D.;Behrens, Timothy E. J.;Rushworth, Matthew F. S.
通讯作者: Rushworth, Matthew F. S.
DOI: 10.1037/0097-7403.21.3.203
发表时间: 1995-07-01
期刊: JOURNAL OF EXPERIMENTAL PSYCHOLOGY-ANIMAL BEHAVIOR PROCESSES
影响因子: --
作者:
BALLEINE, BW;GARNER, C;DICKINSON, A
通讯作者: DICKINSON, A
DOI: 10.1038/nn.3515
发表时间: 2013-10
影响因子: 25
作者:
Barron, Helen C.;Dolan, Raymond J.;Behrens, Timothy E. J.
通讯作者: Behrens, Timothy E. J.
DOI: 10.1001/archpsyc.63.12.1386
发表时间: 2006-12-01
影响因子: --
作者:
Boileau, Isabelle;Dagher, Alain;Benkelfat, Chawki
通讯作者: Benkelfat, Chawki