Theory of Reinforcement Learning and Algorithms of Route Choice in Transportation Networks
Theory of Reinforcement Learning and Algorithms of Route Choice in Transportation Networks
批准号:
22360201
负责人:
MIYAGI Toshihiko
金额:
$5.24万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2010
资助国家:
日本
项目状态:
已结题
起止时间:
2010 至 2012
中文摘要
该研究表明,在交通网络中的个人旅行者被严格建模为自适应学习代理人谁接收的旅行信息,通过日常的日常经验,使他的决定,以加强他的行动依赖于实现的回报。提出了一种与理论相一致的自适应学习算法,并证明了该算法使系统以概率1到达纳什均衡。所提出的算法进行了数值测试,使用的例子网络与各种定义不明确的链路成本函数,并检查算法的快速收敛。此外,我们还提出了一种估计方法的结构参数中包含的路径选择模型。应用于室内实验获得的日常路径选择数据,效果令人满意。
英文摘要
This research shows that an individual traveler in transportation networks is rigorously modeled as an adaptive learning agent who receives travel information through day-to-day experience and makes his decision so as to reinforce his action depending the realized payoffs. An adaptive learning algorithm consistent with the theory is proposed and proved that it leads the system to a Nash equilibrium with probability one. The proposed algorithms have tested numerically by using example networks with various ill-defined link cost functions and examined a rapid convergence of the algorithms. In addition, we have proposed an estimation method for the structure parameters included in the route choice model. The application to the data of theday-to-day route choice obtained by the indoor experiments was satisfactory.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
nformed-user algorithms that converges to Nash equilibrium in traffic games
流量博弈中收敛到纳什均衡的知情用户算法
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[Regmi, R.K., Nakagawa, H., Kawaike, K., Baba, Y., Zhang, H., 重谷祐樹, G.C. Peque,Jr.]
通讯作者:
G.C. Peque,Jr.
カルマンフィルターを応用した所要時間推定法の提案実用性
提出了应用卡尔曼滤波器所需时间估计方法的实用性
DOI:
--
发表时间:
2010
期刊:
影响因子:
--
作者:
[Zhang H., Nakagawa, H. and Mizutani, H, 村上大輔・堤盛人, Mohammad Farid, 高橋雅憲,高山純一,中山晶一朗, 渡部桂子, 宮田輝星・宮城俊彦]
通讯作者:
宮田輝星・宮城俊彦
経路選択行動に関する室内実験
路径选择行为的实验室实验
DOI:
--
发表时间:
2013
期刊:
交通工学
影响因子:
--
作者:
[Ryosuke Arai, So Kazama, Sinji Takahashi and Yasuhiro Takemon, A. Matsumoto, 池田愛,宮城俊彦]
通讯作者:
池田愛,宮城俊彦
社会資本整備を内包した経済成長モデルのパラメータ推定
包括社会资本发展的经济增长模型参数估计
DOI:
--
发表时间:
2010
期刊:
土木計画学研究・論文集
影响因子:
--
作者:
[Toru Hagiwara, Hidekatsu Hamaoka, 加藤裕人・宮城俊彦・仲原由布子]
通讯作者:
加藤裕人・宮城俊彦・仲原由布子
Informed-user algorithms that converges to Nash equilibrium in traffic games
在流量博弈中收敛到纳什均衡的知情用户算法
DOI:
--
发表时间:
2012
期刊:
Procedia-Social and Behavioral Sciences
影响因子:
--
作者:
[Miyagi, T., and G.C. Peque,Jr.]
通讯作者:
and G.C. Peque,Jr.
共 25 条
A Study on Dynamic Traffic Assignment Based on An Atomic Model of Route-Choice
-
批准号:26420511
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$3.24万
-
财政年份:2014
-
负责人:MIYAGI Toshihiko
-
依托单位:
The Study on Development and Applicability of Knowledge-Based Learning Algorithm for Route Guidance
-
批准号:18560519
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.06万
-
财政年份:2006
-
负责人:MIYAGI Toshihiko
-
依托单位:
Non-surveying Construction of a 47 Interregional Input-Output Table and Calibration of SCGE Model
-
批准号:15560458
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.37万
-
财政年份:2003
-
负责人:MIYAGI Toshihiko
-
依托单位:
Sensitivity Analysis for Multiregional General Equilibrium Models
-
批准号:13650582
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.54万
-
财政年份:2001
-
负责人:MIYAGI Toshihiko
-
依托单位:
Integration of Transportation Planning Process Combining with Demand Forecasting Process
-
批准号:11650545
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.92万
-
财政年份:1999
-
负责人:MIYAGI Toshihiko
-
依托单位:
A STUDY ON APPLIED NETWORK EQUILIBRIUM MODELS
-
批准号:07650618
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.34万
-
财政年份:1995
-
负责人:MIYAGI Toshihiko
-
依托单位:
A formulation of spatial price equilibrium model and its computation procedure
-
批准号:63550387
-
项目类别:Grant-in-Aid for General Scientific Research (C)
-
资助金额:$1.34万
-
财政年份:1988
-
负责人:MIYAGI Toshihiko
-
依托单位:
海外基金