The Good, the Bad, and the Irrelevant: Neural Mechanisms of Learning Real and Hypothetical Rewards and Effort

The Good, the Bad, and the Irrelevant: Neural Mechanisms of Learning Real and Hypothetical Rewards and Effort
复制标题

好的、坏的和不相关的:学习真实和假设的奖励和努力的神经机制

DOI:
--
复制
发表时间:
2015
影响因子:
5.3
通讯作者:
M. Rushworth
M. Rushworth
中科院分区:
医学1区
文献类型:
--
作者:
J. Scholl;Nils Kolling;N. Nelissen;Marco K. Wittmann;C. Harmer;M. Rushworth

文献摘要

参考文献

被引文献

相似文献

自然环境是复杂的,单一的选择可能会导致多个结果。代理人应该了解哪些结果是由于他们的选择而导致的,因此与未来的决策相关,以及哪些结果是随机的,对于所有选择来说是共同的,因此与选项之间的未来决策无关。我们设计了一个实验,让人类参与者学习两个选项的不同奖励和努力程度,并在其中重复选择。与选择相关的奖励是随机的、真实的或假设的(即,参与者只有时会收到与所选选项相关的奖励大小)。然而,任何一项试验的奖励的真实/假设性质与了解选择的长期价值无关,参与者应该只关注结果的信息性内容,而不考虑它是真实的还是假设的奖励。然而,我们发现,参与者表现出一种非理性的选择偏见,他们更喜欢在上次试验中偶然导致真正回报的选择。杏仁核和腹内侧前额叶活动与参与者的选择因实际的奖励收受而产生偏差的方式有关。相比之下,背侧前扣带回皮质、额盖/前脑岛,特别是外侧前额叶皮质的活动与参与者抵制这种偏见的程度有关,并以与特定选择具有真实和更持续的关系的结果方面指导有效选择,抑制不相关的奖励信息,以实现更优的学习和决策。在复杂的自然环境中,单一的选择可能会导致多个结果。人类代理人应该只从他们选择的结果中学习,而不是从没有这种关系的结果中学习。我们设计了一项实验,在一个结果的其他特征是随机的、与选择无关的环境中,测量关于奖励和努力程度的学习。我们发现,尽管人们可以了解奖励的大小,但他们仍然非理性地偏向于根据随机奖励特征的存在或不存在而重复某些选择。前额叶皮质不同脑区的活动要么反映了偏见,要么反映了对偏见的抵抗。
Natural environments are complex, and a single choice can lead to multiple outcomes. Agents should learn which outcomes are due to their choices and therefore relevant for future decisions and which are stochastic in ways common to all choices and therefore irrelevant for future decisions between options. We designed an experiment in which human participants learned the varying reward and effort magnitudes of two options and repeatedly chose between them. The reward associated with a choice was randomly real or hypothetical (i.e., participants only sometimes received the reward magnitude associated with the chosen option). The real/hypothetical nature of the reward on any one trial was, however, irrelevant for learning the longer-term values of the choices, and participants ought to have only focused on the informational content of the outcome and disregarded whether it was a real or hypothetical reward. However, we found that participants showed an irrational choice bias, preferring choices that had previously led, by chance, to a real reward in the last trial. Amygdala and ventromedial prefrontal activity was related to the way in which participants' choices were biased by real reward receipt. By contrast, activity in dorsal anterior cingulate cortex, frontal operculum/anterior insula, and especially lateral anterior prefrontal cortex was related to the degree to which participants resisted this bias and chose effectively in a manner guided by aspects of outcomes that had real and more sustained relationships with particular choices, suppressing irrelevant reward information for more optimal learning and decision making. SIGNIFICANCE STATEMENT In complex natural environments, a single choice can lead to multiple outcomes. Human agents should only learn from outcomes that are due to their choices, not from outcomes without such a relationship. We designed an experiment to measure learning about reward and effort magnitudes in an environment in which other features of the outcome were random and had no relationship with choice. We found that, although people could learn about reward magnitudes, they nevertheless were irrationally biased toward repeating certain choices as a function of the presence or absence of random reward features. Activity in different brain regions in the prefrontal cortex either reflected the bias or reflected resistance to the bias.
DOI: 10.1016/j.mri.2007.03.007
发表时间: 2007-12-01
影响因子: 2.5
作者:
Rogers, Baxter P.;Morgan, Victoria L.;Gore, John C.
通讯作者: Gore, John C.