The Resilience of Cooperation in a Dilemma Game Played by Reinforcement Learning Agents
The Resilience of Cooperation in a Dilemma Game Played by Reinforcement Learning Agents
复制标题
强化学习代理在困境博弈中的合作弹性
DOI:
10.1109/agents.2017.8015297
复制
发表时间:
2017
期刊:
影响因子:
--
通讯作者:
and Nobuhiro Inuzuka
中科院分区:
文献类型:
--
作者:
Koichi Moriyama;Kaori Nakase;Atsuko Mutoh;and Nobuhiro Inuzuka
This work discusses what an (independent) reinforcement learning agent can do in a multiagent environment. In particular, we consider a stateless Q-learning agent in a Prisoner's Dilemma (PD) game. Although it had been shown in the literature that stateless, independent Q-learning agents had been difficult to cooperate with each other in an iterated PD (IPD) game, we gave a condition of PD payoffs and Q-learning parameters that helps the agents cooperate with each other. Based on the condition, we also discussed the ratio of mutual cooperation happening in IPD games. It supposed that mutual cooperation was fragile, i.e., one misfortune defection would have the agents slide down the spiral of mutual defection. However, it is not always correct. Mutual cooperation will reinforce itself and thus it will be robust and resilient. Hence, this work analytically derives how long a series of mutual cooperation continues once it happened while considering the resilience. It gives us further comprehension of the process of reinforcement learning in IPD games.