Deconstructing the human algorithms for exploration.
Deconstructing the human algorithms for exploration.
复制标题
DOI:
10.1016/j.cognition.2017.12.014
复制
发表时间:
2018-04
期刊:
影响因子:
3.4
通讯作者:
Gershman SJ
中科院分区:
文献类型:
--
作者:
Gershman SJ
The dilemma between information gathering (exploration) and reward seeking (exploitation) is a fundamental problem for reinforcement learning agents. How humans resolve this dilemma is still an open question, because experiments have provided equivocal evidence about the underlying algorithms used by humans. We show that two families of algorithms can be distinguished in terms of how uncertainty affects exploration. Algorithms based on uncertainty bonuses predict a change in response bias as a function of uncertainty, whereas algorithms based on sampling predict a change in response slope. Two experiments provide evidence for both bias and slope changes, and computational modeling confirms that a hybrid model is the best quantitative account of the data.
登录
查看更多内容
影响因子:
64.8
作者:
Daw, Nathaniel D.;O'Doherty, John P.;Dayan, Peter;Seymour, Ben;Dolan, Raymond J.
通讯作者:
Dolan, Raymond J.
DOI:
10.1098/rstb.2007.2098
发表时间:
2007-05-29
影响因子:
6.3
作者:
Cohen, Jonathan D.;McClure, Samuel M.;Yu, Angela J.
通讯作者:
Yu, Angela J.
影响因子:
1.8
作者:
Gershman, Samuel J.;Tenenbaum, Joshua B.;Jaekel, Frank
通讯作者:
Jaekel, Frank
影响因子:
7.5
作者:
Auer, P;Cesa-Bianchi, N;Fischer, P
通讯作者:
Fischer, P
影响因子:
3.9
作者:
Lee, Michael D.;Zhang, Shunan;Steyvers, Mark
通讯作者:
Steyvers, Mark