Flexible control as surrogate reward or dynamic reward maximization

Flexible control as surrogate reward or dynamic reward maximization
复制标题

灵活控制作为替代奖励或动态奖励最大化

DOI:
10.1016/j.cognition.2022.105262
复制
发表时间:
2022
期刊:
影响因子:
3.4
通讯作者:
Liljeholm, Mimi
Liljeholm, Mimi
中科院分区:
心理学2区
文献类型:
--
作者:
Liljeholm, Mimi

文献摘要

参考文献

相似文献

特定体验的效用,如与特定的朋友互动或品尝特定的食物,根据稳态和享乐原则不断波动。因此,为了使奖励最大化,个体必须能够在偏好发生变化时,通过在不同行为之间切换来逃避或获得结果。最近关于人类和人工智能的研究已经用信息论的术语定义了这种灵活的工具控制,并假设它可以作为奖励的替代品。然而,另一种可能性是,灵活控制所提供的适应性是通过对结果值的动态变化进行规划而默认实现的。在当前的研究中,一种预期的实用新型计算了与感官结果相关的一系列可能的金钱收益和损失的决策值,提供了最适合行为选择数据的方法,并在获得奖励方面表现最好。此外,与先前关于感知控制和人格的研究一致,维度分裂型的个体差异与灵活控制水平最高和最低条件下的行为选择偏好相关。这些结果促进了工具控制在目标导向选择中的作用的文献的增长。
The utility of a given experience, like interacting with a particular friend or tasting a particular food, fluctuates continually according to homeostatic and hedonic principles. Consequently, to maximize reward, an individual must be able to escape or attain outcomes as preferences change, by switching between actions. Recent work on human and artificial intelligence has defined such flexible instrumental control in information theoretic terms and postulated that it may serve as a reward surrogate. Another possibility, however, is that the adaptability afforded by flexible control is tacitly implemented by planning for dynamic changes in outcome values. In the current study, an expected utility model that computes decision values over a range of possible monetary gains and losses associated with sensory outcomes provided the best fit to behavioral choice data and performed best in terms of earned rewards. Moreover, consistent with previous work on perceived control and personality, individual differences in dimensional schizotypy were correlated with behavioral choice preferences in conditions with the greatest and lowest levels of flexible control. These results contribute to a growing literature on the role of instrumental control in goal-directed choice.
DOI: 10.3389/fnbeh.2016.00043
发表时间: 2016
影响因子: 3
作者:
Garbarini F;Mastropasqua A;Sigaudo M;Rabuffetti M;Piedimonte A;Pia L;Rocca P
通讯作者: Rocca P
猴子(Macaca Fascillaryis)的强迫选择和自由选择
DOI: 10.2466/pms.1999.88.1.242
发表时间: 1999
影响因子: 1.6
作者:
Shuji Suzuki
通讯作者: Shuji Suzuki
选择作为一种价值
DOI: 10.2466/pr0.1970.26.3.912
发表时间: 1970
影响因子: 2.3
作者:
S. C. Voss;M. Homzie
通讯作者: M. Homzie
DOI: 10.1523/jneurosci.2463-19.2020
发表时间: 2020
期刊: The Journal of Neuroscience
影响因子: --
作者:
Norton, Kaitlyn G.;Liljeholm, Mimi
通讯作者: Liljeholm, Mimi
DOI: 10.1038/srep36295
发表时间: 2016-11-04
期刊: Scientific reports
影响因子: 4.6
作者:
Mistry P;Liljeholm M
通讯作者: Liljeholm M