Dynamic decision policy reconfiguration under outcome uncertainty.

Dynamic decision policy reconfiguration under outcome uncertainty.
复制标题

DOI:
10.7554/elife.65540
复制
发表时间:
2021-12-24
期刊:
影响因子:
7.7
通讯作者:
Verstynen T
Verstynen T
中科院分区:
生物学1区
文献类型:
--
作者:
Bond K;Dunovan K;Porter A;Rubin JE;Verstynen T

文献摘要

参考文献

被引文献

相似文献

在不确定或不稳定的环境中,有时最好的决定是改变主意。为了阐明这种灵活性,我们评估了当最有价值的行动发生变化时,底层决策策略如何适应。人类参与者执行了一个动态的双臂强盗任务,操纵相对奖励的确定性(冲突)和行动结果的可靠性(波动性)。冲突和波动性的连续估计,通过改变证据积累的速度(漂移率)和作出决定所需的证据量(边界高度),分别在探索状态的变化。在试验水平上,在最佳选择的切换之后,漂移率骤降,边界高度微弱地尖峰,导致缓慢的探索状态。我们发现,漂移率驱动大部分的这种响应,在整个实验的边界高度的不可靠的贡献。令人惊讶的是,我们没有发现任何证据表明瞳孔反应与决策政策的变化。我们的结论是,人类在应对环境变化的决策政策中表现出一种刻板的转变。
In uncertain or unstable environments, sometimes the best decision is to change your mind. To shed light on this flexibility, we evaluated how the underlying decision policy adapts when the most rewarding action changes. Human participants performed a dynamic two-armed bandit task that manipulated the certainty in relative reward (conflict) and the reliability of action-outcomes (volatility). Continuous estimates of conflict and volatility contributed to shifts in exploratory states by changing both the rate of evidence accumulation (drift rate) and the amount of evidence needed to make a decision (boundary height), respectively. At the trialwise level, following a switch in the optimal choice, the drift rate plummets and the boundary height weakly spikes, leading to a slow exploratory state. We find that the drift rate drives most of this response, with an unreliable contribution of boundary height across experiments. Surprisingly, we find no evidence that pupillary responses associated with decision policy changes. We conclude that humans show a stereotypical shift in their decision policies in response to environmental changes.
DOI: 10.1371/journal.pcbi.1006033
发表时间: 2018-04
影响因子: 4.3
作者:
Caballero JA;Humphries MD;Gurney KN
通讯作者: Gurney KN