Study on Decentralized Learning Algorithms in Markovian Environments
Study on Decentralized Learning Algorithms in Markovian Environments
批准号:
06650449
负责人:
ABE Kenichi
金额:
$1.34万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for General Scientific Research (C)
财政年份:
1994
资助国家:
日本
项目状态:
已结题
起止时间:
1994 至 1995
中文摘要
本文的主要研究成果如下:(1)针对动态未知的马尔可夫链,提出了一种新的分散学习算法。详细的仿真研究揭示了我们算法的可行性及其相对于Q学习方案的优越性。(2)我们提出了一个面向对象的设计支持系统开发自主移动的机器人。我们的支持系统的有用性进行了检查,通过一些模拟和真实的机器人的实现。通过使用这个支持系统,我们还开发了一个机器人,它可以获得一个适当的设置中的避障算法,称为VFH的增益因子,通过学习。(3)将分散学习算法应用于智能移动的机器人的自适应动作选择问题。我们使用了一个机器人,它有两个光电传感器来测量左右方向的光强。机器人的任务是学习一个动作选择策略,用于移动到放置在房间任何位置的灯。分散式学习方法通过运行模拟机器人成功地进行了测试,这些机器人是由面向对象的设计支持系统实现的。(4)提出了一类称为霍隆网络的递阶系统作为非线性动力系统辨识的一般模型。霍隆网络能够通过自组织其结构来进化,并且能够在不假设对它们有太多知识的情况下学习非线性系统。
英文摘要
The results of this study are summarized as follows :(1) We proposed a new decentralized learning algorithm for Markov chains with unknown dynamics. A detailed simulation study revealed the feasibility of our algorithm and its superiority to the Q-learning scheme.(2) We proposed an object-oriented design support system for developing autonomous mobile robots. The usefulness of our support system was examined through some implementations of simulated and real robots. By using this support system, we also developed a robot which can acquire a proper setting of gain factors in an obstacle avoidance algorithm, called the VFH,by learning.(3) We applied the decentralized learning algorithm to the problem of adaptive action selection in an intelligent mobile robot. We employed a robot which has two photosensors to measure the light intensity in right or left direction. The robot's task is to learn an action selection policy for moving toward and getting to a light placed in any location of a room. The decentralized learning approach was successfully tested by running simulated robots, which were implemented by the object-oriented design support system.(4) We proposed a class of hierarchical systems called holon networks as general models for identification of nonlinear dynamical systems. Holon networks are able to evolve by self-organizing their structure and learn nonlinear systems without assuming much knowledge of them.
期刊论文(42)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
釜谷博行: "学習オートマトンによる移動ロボットナビゲ-タのパラメータ自動調整" 電気学会論文誌. 115-C. 1570-1571 (1995)
Hiroyuki Kamaya:“使用学习自动机自动调整移动机器人导航器”,日本电气工程师学会汇刊 115-C 1570-1571 (1995)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
N.Honma: "On Autonomous Decentralized Evolution of Holon Network" Proc.of The 9th KACC Int'l Session. 498-503 (1994)
N.Honma:“论 Holon 网络的自主去中心化演化”第 9 届 KACC 国际会议议程。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
H.Honma: "Adaptive Evolution of Holon Networks by an Autonomous Decentralizes Method" International Symposium on Artificial Life. (1996)
H.Honma:“通过自主分散方法实现 Holon 网络的自适应进化”国际人工生命研讨会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
本間経康: "自律分散的適応制御によるホロンネットワークの進化について" 計測自動制御学会論文集. 31(印刷中). (1995)
Tsuneyasu Honma:“论通过自主分散自适应控制的全子网络的演化”,仪器与控制工程师学会汇刊 31(出版中)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
釜谷 博行: "オブジェクト指向設計に基づいた自律型移動ロボットの開発支援システム" 電気学会論文誌. 115-C. 819-828 (1995)
Hiroyuki Kamaya:“基于面向对象设计的自主移动机器人的开发支持系统”日本电气工程师学会汇刊 115-C 819-828(1995)。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 21 条
Studies on Literary History in Bohemia
-
批准号:19K00493
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.75万
-
财政年份:2019
-
负责人:ABE Kenichi
-
依托单位:
Studies on Images of "East" in East European Literature
-
批准号:24320064
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$6.99万
-
财政年份:2012
-
负责人:ABE Kenichi
-
依托单位:
Self-Organization of Hierarchical Reinforcement Learning System
-
批准号:13650480
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.18万
-
财政年份:2001
-
负责人:ABE Kenichi
-
依托单位:
Self-control of Memory Structure of Reinforcement Learning in Hidden Markov Environments
-
批准号:11650441
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$2.24万
-
财政年份:1999
-
负责人:ABE Kenichi
-
依托单位:
Study on Decentralized Learning Algorithms in Non-Markovian Environments
-
批准号:09650451
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$1.6万
-
财政年份:1997
-
负责人:ABE Kenichi
-
依托单位:
海外基金