Scaling Genetic Programming to Complex Reinforcement Learning Tasks
Scaling Genetic Programming to Complex Reinforcement Learning Tasks
批准号:
RGPIN-2020-04438
负责人:
Heywood, Malcolm
金额:
$2.11万
依托单位:
依托单位国家:
加拿大
项目类别:
Discovery Grants Program - Individual
财政年份:
2020
资助国家:
加拿大
项目状态:
已结题
起止时间:
2020-01-01 至 2021-12-31
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Reinforcement learning (RL) represents a type of task in which an agent interacts with an environment to maximize its long term reward. A lot of progress has recently been made with deep learning under high-dimensional state and action spaces. This means that rather than having to first develop a suite of appropriate input features, sensors such as video can be used directly. An enormous number of applications have benefited from this development, from algorithms that play Go and Chess better than humans, to facilitating new levels of human competitive performance for robot control tasks. However, one drawback of such an approach is that they generally represent complex black box solutions that require hardware support to deploy, even after training.
We recently proposed an alternative approach for scaling RL to high-dimensional state spaces using genetic programming. To do so, teams of programs self organize into Tangled Program Graphs (TPG), which represents an approach of organizing teams of programs into graphs. Our initial benchmarking under high-dimensional RL tasks demonstrates that equivalent quality solutions can be discovered, but with multiple orders of magnitude lower complexity.
The proposed research program will greatly expand on the TPG approach to efficiently discover solutions to non-reactive RL tasks requiring multiple simultaneous actions per time step. The long term research program is organized around three objectives:
1) Support for the Automatic identification of behavioural subgraphs: provides the basis for task transfer, accelerated training and increased transparency of machine learning solutions.
2) Develop Multiple concurrent memory models: is the basis for scaling TPG to a wide cross section of non-reactive RL tasks. Without this, it would not be possible to scale to partially observable problems, a class of tasks of widespread impact.
3) Support for describing actions as Multi-dimensional spaces: means that decisions involving multiple real and discrete actions per state can be made simultaneously. A capability that also potentially appears in many applications.
Successful completion of this research program will result in a TPG framework that provides solution quality complementing those from deep learning. However, TPG constructs solutions by explicitly discovering mechanisms for decomposing the decision making task. This means that solutions are light-weight, executing in real-time without any form of hardware support. The simplicity of solutions will also support insights into attribute support and solution transparency. This is particularly important when attempting to gain knowledge from solutions post training. Success in the proposed research program would demonstrate new models for addressing open ended questions regarding the application and deployment of RL agents to navigation, motor control and strategic decision making in real-time partially observable environments.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Scaling Genetic Programming to Complex Reinforcement Learning Tasks
-
批准号:RGPIN-2020-04438
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$2.11万
-
财政年份:2022
-
负责人:Heywood, Malcolm
-
依托单位:
Scaling Genetic Programming to Complex Reinforcement Learning Tasks
-
批准号:RGPIN-2020-04438
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$2.11万
-
财政年份:2021
-
负责人:Heywood, Malcolm
-
依托单位:
Permutation based task transfer for genetic programming
-
批准号:RGPIN-2015-06117
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.31万
-
财政年份:2019
-
负责人:Heywood, Malcolm
-
依托单位:
Coevolutionary automatic game content generation of physics and flighting style games
-
批准号:499792-2016
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$5.34万
-
财政年份:2018
-
负责人:Heywood, Malcolm
-
依托单位:
Permutation based task transfer for genetic programming
-
批准号:RGPIN-2015-06117
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.31万
-
财政年份:2018
-
负责人:Heywood, Malcolm
-
依托单位:
Permutation based task transfer for genetic programming
-
批准号:RGPIN-2015-06117
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.31万
-
财政年份:2017
-
负责人:Heywood, Malcolm
-
依托单位:
Coevolutionary automatic game content generation of physics and flighting style games
-
批准号:499792-2016
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$5.34万
-
财政年份:2017
-
负责人:Heywood, Malcolm
-
依托单位:
Permutation based task transfer for genetic programming
-
批准号:RGPIN-2015-06117
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.31万
-
财政年份:2016
-
负责人:Heywood, Malcolm
-
依托单位:
Coevolutionary automatic game content generation of physics and flighting style games
-
批准号:499792-2016
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$5.34万
-
财政年份:2016
-
负责人:Heywood, Malcolm
-
依托单位:
Constructing risk predictors for mobile device behaviour analytics
-
批准号:485070-2015
-
项目类别:Engage Grants Program
-
资助金额:$1.82万
-
财政年份:2015
-
负责人:Heywood, Malcolm
-
依托单位:
Evolving under tasks of incomplete information: streaming and self play
-
批准号:451239-2013
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$6.41万
-
财政年份:2015
-
负责人:Heywood, Malcolm
-
依托单位:
Permutation based task transfer for genetic programming
-
批准号:RGPIN-2015-06117
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.31万
-
财政年份:2015
-
负责人:Heywood, Malcolm
-
依托单位:
Game Server Network Analysis Engine
-
批准号:489094-2015
-
项目类别:Engage Grants Program
-
资助金额:$1.82万
-
财政年份:2015
-
负责人:Heywood, Malcolm
-
依托单位:
EEG artifact removal under minimal sensor redundancy
-
批准号:471475-2014
-
项目类别:Engage Grants Program
-
资助金额:$1.82万
-
财政年份:2014
-
负责人:Heywood, Malcolm
-
依托单位:
Continuous symbiotic program evolution
-
批准号:238791-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.82万
-
财政年份:2014
-
负责人:Heywood, Malcolm
-
依托单位:
Evolving under tasks of incomplete information: streaming and self play
-
批准号:451239-2013
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$6.41万
-
财政年份:2014
-
负责人:Heywood, Malcolm
-
依托单位:
Continuous symbiotic program evolution
-
批准号:238791-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.82万
-
财政年份:2013
-
负责人:Heywood, Malcolm
-
依托单位:
Evolving under tasks of incomplete information: streaming and self play
-
批准号:451239-2013
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$6.41万
-
财政年份:2013
-
负责人:Heywood, Malcolm
-
依托单位:
Continuous symbiotic program evolution
-
批准号:238791-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.82万
-
财政年份:2012
-
负责人:Heywood, Malcolm
-
依托单位:
Pattern Validation in Video Lottery Gaming
-
批准号:408123-2010
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$6.81万
-
财政年份:2011
-
负责人:Heywood, Malcolm
-
依托单位:
海外基金