Retrospective Revaluation in Sequential Decision Making: A Tale of Two Systems

Retrospective Revaluation in Sequential Decision Making: A Tale of Two Systems
复制标题

DOI:
10.1037/a0030844
复制
发表时间:
2014-02-01
影响因子:
4.1
通讯作者:
Otto, A. Ross
Otto, A. Ross
中科院分区:
心理学1区
文献类型:
--
作者:
Gershman, Samuel J.;Markman, Arthur B.;Otto, A. Ross

文献摘要

被引文献

相似文献

最近的人类和动物决策计算理论描绘了两个陷入行为控制之战的系统。一种系统(各种称为无模型或习惯性偏爱先前导致奖励的行为),而第二种称为基于模型或目标导向的系统,偏爱根据代理的内部环境模型因果导致奖励的行为。一些证据表明,可以使用神经或行为操纵在这些系统之间转移控制权,但其他证据表明,这些系统比竞争帐户所暗示的更加相互交织。在 4 个行为实验中,使用回顾性重估设计和认知负荷操纵,我们表明人类决策与合作架构更加一致,其中无模型系统控制行为,而基于模型的系统通过重放和模拟经验来训练无模型系统。
Recent computational theories of decision making in humans and animals have portrayed 2 systems locked in a battle for control of behavior. One system-variously termed model-free or habitual-favors actions that have previously led to reward, whereas a second-called the model-based or goal-directed system-favors actions that causally lead to reward according to the agent's internal model of the environment. Some evidence suggests that control can be shifted between these systems using neural or behavioral manipulations, but other evidence suggests that the systems are more intertwined than a competitive account would imply. In 4 behavioral experiments, using a retrospective revaluation design and a cognitive load manipulation, we show that human decisions are more consistent with a cooperative architecture in which the model-free system controls behavior, whereas the model-based system trains the model-free system by replaying and simulating experience.