Learning to Cooperate in Repeated Games
Learning to Cooperate in Repeated Games
批准号:
9602082
负责人:
In-Koo Cho
金额:
$18.2万
依托单位:
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
1996
资助国家:
美国
项目状态:
已结题
起止时间:
1996-08-01 至 1998-11-09
中文摘要
研究重复博弈的一个主要原因是了解自私的参与者如何在没有共谋协议的情况下协调他们的行动以实现改进。 不幸的是,现有的博弈论模型承认如此多的结果,以至于不可能预测是否会出现协调。 分析还假设一个理性的代理人谁拥有无限的计算能力和完美的远见。 这些假设对均衡模型来说是至关重要的,但也相当不现实。 该项目探讨了替代模型,其中完全理性的代理人被“有限理性”代理人所取代,这些代理人只有有限的计算能力,并且无法完全预见其他玩家的策略,他们必须从过去的经验中学习。 这种方法被证明是捕捉学习动态,并允许应用程序广泛的重复和动态的游戏,有一个“大”的球员谁可以影响长期运行的结果模型。 申请包括国际债务 和道德风险下的最优增长 更具体地说,该项目研究了两个人的重复游戏,其中每个玩家根据梯度法学习对手的策略,假设对手是根据线性策略进行游戏。 此外,每个玩家都人为地添加了随机噪音,这些噪音会慢慢消失,以便对对手的策略进行实验。 对可行策略没有限制,但每个参与者的预测必须是过去观察的线性函数。 选择这类特殊策略的原因是这些策略足够简单,可以很容易地参数化。 然后,每个玩家可以使用最小二乘估计来学习对手的策略。 代理人的偏好也略有修改,使他选择最佳对策,同时最大限度地减少决策过程的复杂性。 然后,得到递归最小二乘学习模型,其中每个参与者随着游戏的进行更新他的信念以及他的操作游戏策略。 学习动态以概率1收敛,并且在极限下,两个参与者具有相同的估计量。 因此,两个参与者的行为是高度相关的,并且结果的极限频率可以通过线性策略中的某些纳什均衡来维持。 例如,在囚徒困境博弈中,结果的极限频率必须是合作和背叛的严格凸组合,这意味着参与者必须学会以正概率合作。
英文摘要
A primary reason for studying repeated games is to understand how selfish players can coordinate their actions to achieve improvements without a collusive agreement. Unfortunately existing game-theoretic models admit so many outcomes that it is impossible to predict whether coordination will emerge. Also analysis postulates a rational agent who has unbounded computational capability and perfect foresight. These assumptions are critical for equilibrium models but also rather unrealistic. This project explores alternative models in which perfectly rational agents are replaced by `boundedly rational` agents who have only limited computational capabilities, and who cannot perfectly foresee the strategy of other players, which they have to learn from the past experiences. This approach is shown to capture learning dynamics and to permit applications to a wide class of repeated and dynamic games which have a `big` player who can influence the long run outcome of the model. Applications include international debt and optimal growth with moral hazard. More specifically, the project examines two person repeated games where each player learns the opponent's strategy according to the gradient method by assuming that the opponent is playing according to a linear strategy. In addition, each player artificially adds random noise that disappears slowly in order to experiment against the opponent's strategy. No restrictions are imposed on feasible strategies, but the forecast of each player must be a linear function of past observations. The reason for selecting this particular class of strategies is that these strategies are simple enough to be parameterized easily. Then, each player can learn the opponent's strategy using least squares estimation. The agent's preference is also modified slightly so that he is selecting a best response while minimizing the complexity of the decision making process. then, a recursive least squares learning model is obtained, where each player updates his belief as well as his operated game strategy as the game proceeds. The learning dynamics converges with probability 1 and in the limit, both players have an identical estimator. Consequently the behavior of the two players is highly correlated, and the limit frequency of outcomes can be sustained by some Nash equilibrium in linear strategies. In the prisoner's dilemma game, for example, the limit frequency of outcomes must be a strict convex combination of cooperation and defection, which implies that the players must learn to cooperative with positive probability.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Machine Learning in Macroeconomic Modeling
-
批准号:1952882
-
项目类别:Standard Grant
-
资助金额:$25.21万
-
财政年份:2019
-
负责人:In-Koo Cho
-
依托单位:
Learning with Model Uncertainty and Misspecification
-
批准号:1952874
-
项目类别:Standard Grant
-
资助金额:$6.26万
-
财政年份:2019
-
负责人:In-Koo Cho
-
依托单位:
Machine Learning in Macroeconomic Modeling
-
批准号:1824253
-
项目类别:Standard Grant
-
资助金额:$31.1万
-
财政年份:2018
-
负责人:In-Koo Cho
-
依托单位:
Learning with Model Uncertainty and Misspecification
-
批准号:1530589
-
项目类别:Standard Grant
-
资助金额:$33.47万
-
财政年份:2015
-
负责人:In-Koo Cho
-
依托单位:
Social Foundation of Nash Bargaining Solution
-
批准号:1061855
-
项目类别:Standard Grant
-
资助金额:$27.68万
-
财政年份:2011
-
负责人:In-Koo Cho
-
依托单位:
Studies on Dynamic Markets: Small Change and Big Impact
-
批准号:0720592
-
项目类别:Standard Grant
-
资助金额:$14.67万
-
财政年份:2007
-
负责人:In-Koo Cho
-
依托单位:
Learning With Misspecified Models
-
批准号:0004315
-
项目类别:Continuing Grant
-
资助金额:$23.09万
-
财政年份:2001
-
负责人:In-Koo Cho
-
依托单位:
Learning to Cooperate in Repeated Games
-
批准号:9996058
-
项目类别:Continuing Grant
-
资助金额:$4.5万
-
财政年份:1998
-
负责人:In-Koo Cho
-
依托单位:
Perceptrons Play Repeated Games: New Approach to Bounded Rationality
-
批准号:9596161
-
项目类别:Standard Grant
-
资助金额:$4.28万
-
财政年份:1995
-
负责人:In-Koo Cho
-
依托单位:
Perceptrons Play Repeated Games: New Approach to Bounded Rationality
-
批准号:9223483
-
项目类别:Standard Grant
-
资助金额:$10.15万
-
财政年份:1993
-
负责人:In-Koo Cho
-
依托单位:
Learning in Dynamic Games
-
批准号:9022642
-
项目类别:Standard Grant
-
资助金额:$5.55万
-
财政年份:1991
-
负责人:In-Koo Cho
-
依托单位:
Uncertainty and Delay in Bargaining
-
批准号:8618596
-
项目类别:Standard Grant
-
资助金额:$10.78万
-
财政年份:1987
-
负责人:In-Koo Cho
-
依托单位:
海外基金