Understanding dopamine and reinforcement learning: The dopamine reward prediction error hypothesis

Understanding dopamine and reinforcement learning: The dopamine reward prediction error hypothesis
复制标题

DOI:
10.1073/pnas.1014269108
复制
发表时间:
2011-09-13
影响因子:
11.1
通讯作者:
Glimcher, Paul W.
Glimcher, Paul W.
中科院分区:
综合性期刊1区
文献类型:
--
作者:
Glimcher, Paul W.

文献摘要

被引文献

相似文献

中脑多巴胺能神经元的研究取得了一些新的进展。理解这些进展以及它们之间的关系需要深入理解作为解释框架并指导正在进行的实验研究的计算模型。理论和实验的结合清楚地表明,中脑多巴胺神经元的阶段性活动为突触修饰提供了一种全局机制。反过来,这些突触修饰为一类特定的强化学习机制提供了机械基础,这些机制现在似乎是人类和动物行为的基础。这篇评论描述了这一结论的根本的关键实证研究结果和得出这一结论的奇妙的理论进展。
A number of recent advances have been achieved in the study of midbrain dopaminergic neurons. Understanding these advances and how they relate to one another requires a deep understanding of the computational models that serve as an explanatory framework and guide ongoing experimental inquiry. This intertwining of theory and experiment now suggests very clearly that the phasic activity of the midbrain dopamine neurons provides a global mechanism for synaptic modification. These synaptic modifications, in turn, provide the mechanistic underpinning for a specific class of reinforcement learning mechanisms that now seem to underlie much of human and animal behavior. This review describes both the critical empirical findings that are at the root of this conclusion and the fantastic theoretical advances from which this conclusion is drawn.