Spatiotemporal neural characterization of prediction error valence and surprise during reward learning in humans.

Spatiotemporal neural characterization of prediction error valence and surprise during reward learning in humans.
复制标题

DOI:
10.1038/s41598-017-04507-w
复制
发表时间:
2017-07-06
期刊:
影响因子:
4.6
通讯作者:
Philiastides MG
Philiastides MG
中科院分区:
综合性期刊3区
文献类型:
--
作者:
Fouragnan E;Queirazza F;Retzler C;Mullinger KJ;Philiastides MG

文献摘要

参考文献

被引文献

相似文献

奖励学习依赖于奖励与潜在选择的准确关联。这些关联可以通过强化学习机制来实现,该机制使用奖励预测误差(RPE)信号(实际奖励和预期奖励之间的差异)来更新未来的奖励预期。尽管有大量关于RPE对学习的影响的文献,但很少有人研究RPE效价(阳性或阴性)和惊喜(偏离预期的绝对程度)的潜在单独贡献。在这里,我们耦合单次试验脑电图与同时获得的功能磁共振成像,在一个概率反向学习任务,提供证据的时间重叠,但在很大程度上不同的空间表示的RPE效价和惊喜。RPE效价的电生理变化与人类奖励网络促进接近或回避学习区域的活动相关。RPE惊喜的电生理变化主要与控制学习速度的人类注意力网络区域的活动相关。至关重要的是,尽管这些表征在很大程度上是独立的空间延伸,但我们的EEG信息功能磁共振成像方法独特地揭示了两个RPE组件在一个较小的网络中的线性叠加,该网络包括视觉记忆和奖励区域。该网络中的活动进一步预测刺激值更新,表明两种信号对奖励学习的贡献相当。
Reward learning depends on accurate reward associations with potential choices. These associations can be attained with reinforcement learning mechanisms using a reward prediction error (RPE) signal (the difference between actual and expected rewards) for updating future reward expectations. Despite an extensive body of literature on the influence of RPE on learning, little has been done to investigate the potentially separate contributions of RPE valence (positive or negative) and surprise (absolute degree of deviation from expectations). Here, we coupled single-trial electroencephalography with simultaneously acquired fMRI, during a probabilistic reversal-learning task, to offer evidence of temporally overlapping but largely distinct spatial representations of RPE valence and surprise. Electrophysiological variability in RPE valence correlated with activity in regions of the human reward network promoting approach or avoidance learning. Electrophysiological variability in RPE surprise correlated primarily with activity in regions of the human attentional network controlling the speed of learning. Crucially, despite the largely separate spatial extend of these representations our EEG-informed fMRI approach uniquely revealed a linear superposition of the two RPE components in a smaller network encompassing visuo-mnemonic and reward areas. Activity in this network was further predictive of stimulus value updating indicating a comparable contribution of both signals to reward learning.
DOI: 10.1523/jneurosci.3793-11.2011
发表时间: 2011-12-07
期刊: The Journal of neuroscience : the official journal of the Society for Neuroscience
影响因子: --
作者:
Asaad WF;Eskandar EN
通讯作者: Eskandar EN
DOI: 10.1016/j.neuron.2015.08.018
发表时间: 2015-09-02
期刊: Neuron
影响因子: 16.2
作者:
Chau BK;Sallet J;Papageorgiou GK;Noonan MP;Bell AH;Walton ME;Rushworth MF
通讯作者: Rushworth MF
DOI: 10.1038/ncomms9107
发表时间: 2015-09-08
影响因子: 16.6
作者:
Fouragnan E;Retzler C;Mullinger K;Philiastides MG
通讯作者: Philiastides MG
DOI: 10.1038/nature04766
发表时间: 2006-06-15
期刊: NATURE
影响因子: 64.8
作者:
Daw, Nathaniel D.;O'Doherty, John P.;Dayan, Peter;Seymour, Ben;Dolan, Raymond J.
通讯作者: Dolan, Raymond J.
DOI: 10.1006/nimg.1999.0498
发表时间: 1999-11-01
期刊: NEUROIMAGE
影响因子: 5.7
作者:
Friston, KJ;Zarahn, E;Dale, AM
通讯作者: Dale, AM