Leading a Best-Response Teammate in an Ad Hoc Team
Leading a Best-Response Teammate in an Ad Hoc Team
复制标题
DOI:
10.1007/978-3-642-15117-0_10
复制
发表时间:
2009-05
期刊:
影响因子:
--
通讯作者:
P. Stone;G. Kaminka;J. Rosenschein
中科院分区:
文献类型:
--
作者:
P. Stone;G. Kaminka;J. Rosenschein
Teams of agents may not always be developed in a planned, coordinated fashion. Rather, as deployed agents become more common in e-commerce and other settings, there are increasing opportunities for previously unacquainted agents to cooperate in ad hoc team settings. In such scenarios, it is useful for individual agents to be able to collaborate with a wide variety of possible teammates under the philosophy that not all agents are fully rational. This paper considers an agent that is to interact repeatedly with a teammate that will adapt to this interaction in a particular suboptimal, but natural way. We formalize this setting in game-theoretic terms, provide and analyze a fully-implemented algorithm for finding optimal action sequences, prove some theoretical results pertaining to the lengths of these action sequences, and provide empirical results pertaining to the prevalence of our problem of interest in random interaction settings.