Multiple systems in macaques for tracking prediction errors and other types of surprise.
Multiple systems in macaques for tracking prediction errors and other types of surprise.
复制标题
DOI:
10.1371/journal.pbio.3000899
复制
发表时间:
2020-10
期刊:
影响因子:
9.8
通讯作者:
Rushworth MFS
中科院分区:
文献类型:
--
作者:
Grohn J;Schüffelgen U;Neubert FX;Bongioanni A;Verhagen L;Sallet J;Kolling N;Rushworth MFS
Animals learn from the past to make predictions. These predictions are adjusted after prediction errors, i.e., after surprising events. Generally, most reward prediction errors models learn the average expected amount of reward. However, here we demonstrate the existence of distinct mechanisms for detecting other types of surprising events. Six macaques learned to respond to visual stimuli to receive varying amounts of juice rewards. Most trials ended with the delivery of either 1 or 3 juice drops so that animals learned to expect 2 juice drops on average even though instances of precisely 2 drops were rare. To encourage learning, we also included sessions during which the ratio between 1 and 3 drops changed. Additionally, in all sessions, the stimulus sometimes appeared in an unexpected location. Thus, 3 types of surprising events could occur: reward amount surprise (i.e., a scalar reward prediction error), rare reward surprise, and visuospatial surprise. Importantly, we can dissociate scalar reward prediction errors—rewards that deviated from the average reward amount expected—and rare reward events—rewards that accorded with the average reward expectation but that rarely occurred. We linked each type of surprise to a distinct pattern of neural activity using functional magnetic resonance imaging. Activity in the vicinity of the dopaminergic midbrain only reflected surprise about the amount of reward. Lateral prefrontal cortex had a more general role in detecting surprising events. Posterior lateral orbitofrontal cortex specifically detected rare reward events regardless of whether they followed average reward amount expectations, but only in learnable reward environments. Animals learn from the past to make predictions, and predictions are adjusted whenever surprising outcomes violate those predictions. This study shows that macaques use multiple systems to detect different types of rare reward events; while activity in the vicinity of the dopaminergic midbrain reflected scalar reward prediction errors (i.e. reward amount), activity in the orbitofrontal cortex reflected a rare reward signal.
登录
查看更多内容
影响因子:
16.2
作者:
Boorman ED;Rajendran VG;O'Reilly JX;Behrens TE
通讯作者:
Behrens TE
影响因子:
3.7
作者:
Caspari, Natalie;Arsenault, John T.;Vanduffel, Wim
通讯作者:
Vanduffel, Wim
影响因子:
16.2
作者:
Daw ND;Gershman SJ;Seymour B;Dayan P;Dolan RJ
通讯作者:
Dolan RJ
影响因子:
5.3
作者:
Hart, Andrew S.;Rutledge, Robb B.;Phillips, Paul E. M.
通讯作者:
Phillips, Paul E. M.
影响因子:
64.8
作者:
Burke, Kathryn A.;Franz, Theresa M.;Schoenbaum, Geoffrey
通讯作者:
Schoenbaum, Geoffrey