Learning good representations for and with reinforcement learning
Learning good representations for and with reinforcement learning
批准号:
RGPIN-2017-06788
负责人:
Precup, Doina
金额:
$5.17万
依托单位:
依托单位国家:
加拿大
项目类别:
Discovery Grants Program - Individual
财政年份:
2019
资助国家:
加拿大
项目状态:
已结题
起止时间:
2019-01-01 至 2020-12-31
中文摘要
人工智能(AI)在隔离智能的不同方面取得了巨大进展,并提出了灵活的表示和强大的算法,从而能够胜任特定任务。例如,AI智能体比人类更擅长玩围棋等游戏,这曾经被认为是不可能的。然而,人类甚至动物通常表现出的那种灵活、健壮和自主的能力仍然难以捉摸。最好的人工智能系统仍然是针对特定问题而调整的。我们的主要研究目标是开发通用的人工智能方法,其核心是强化学习。强化学习是一种从与环境的交互中学习的方法,受到动物学习理论的启发。该提案旨在设计可以自动为强化学习代理创建表示的算法,这些代理允许它们对世界进行建模并在多个时间尺度上采取行动。我们的目标是提供新的优化标准,正式描述什么是一个很好的一套抽象的表示,提供基于梯度的学习算法来学习这样的模型,并证明其有效性,通过经验评估模拟域,游戏,以及真实的时间序列预测数据集。 我们将通过解释智能体应如何在其环境中移动以优化其学习速度来解决探索的关键问题。最后,我们将在其他算法中利用这些方法,这些算法可以从多个时间尺度中受益,例如深度递归神经网络的训练。
英文摘要
Artificial intelligence (AI) has made great progress in isolating different aspects of intelligence and proposing flexible representations and powerful algorithms that lead to competence in specific tasks. For example, AI agents are better than humans at playing games like Go, a feat once considered impossible. However, the sort of flexible, robust, and autonomous competence routinely exhibited by humans, or even animals, remains elusive. The best AI systems are still tuned to specific problems. Our main research goal is to develop general AI methodology that relies, at its core, on reinforcement learning. Reinforcement learning is an approach to learning from interaction with an environment, inspired by animal learning theory. This proposal aims to design algorithms that can automatically create representations for reinforcement learning agents which allow them to model the world and to act at multiple time scales. We aim to provide new optimization criteria which describe formally what is a good set of abstract representations, provide gradient-based learning algorithms to learn such models, and demonstrate their effectiveness through empirical evaluations in simulated domains, game playing, as well as real time series prediction data sets. We will tackle the crucial problem of exploration, by explaining how an agent should move about its environment in order to optimize its learning speed. Finally, we will leverage these methods inside other algorithms that can benefit from multiple time scales, such as the training of deep, recurrent neural networks.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Learning good representations for and with reinforcement learning
-
批准号:RGPIN-2017-06788
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$10.34万
-
财政年份:2021
-
负责人:Precup, Doina
-
依托单位:
Learning good representations for and with reinforcement learning
-
批准号:RGPIN-2017-06788
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$5.17万
-
财政年份:2020
-
负责人:Precup, Doina
-
依托单位:
Learning good representations for and with reinforcement learning
-
批准号:RGPIN-2017-06788
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$5.17万
-
财政年份:2018
-
负责人:Precup, Doina
-
依托单位:
Machine Learning
-
批准号:1000231167-2015
-
项目类别:Canada Research Chairs
-
资助金额:$7.29万
-
财政年份:2017
-
负责人:Precup, Doina
-
依托单位:
Learning good representations for and with reinforcement learning
-
批准号:RGPIN-2017-06788
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$5.17万
-
财政年份:2017
-
负责人:Precup, Doina
-
依托单位:
Machine Learning
-
批准号:1000231167-2015
-
项目类别:Canada Research Chairs
-
资助金额:$14.57万
-
财政年份:2016
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:238988-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$4.37万
-
财政年份:2016
-
负责人:Precup, Doina
-
依托单位:
McGill Science for a Sustainable Society Symposium
-
批准号:490803-2015
-
项目类别:Regional Office Discretionary Funds
-
资助金额:$0.36万
-
财政年份:2015
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:238988-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$4.37万
-
财政年份:2015
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:238988-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$4.37万
-
财政年份:2014
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:238988-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$4.37万
-
财政年份:2013
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:238988-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$4.37万
-
财政年份:2012
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:401375-2010
-
项目类别:Discovery Grants Program - Accelerator Supplements
-
资助金额:$2.91万
-
财政年份:2012
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:401375-2010
-
项目类别:Discovery Grants Program - Accelerator Supplements
-
资助金额:$2.91万
-
财政年份:2011
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:238988-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$4.37万
-
财政年份:2011
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:401375-2010
-
项目类别:Discovery Grants Program - Accelerator Supplements
-
资助金额:$2.91万
-
财政年份:2010
-
负责人:Precup, Doina
-
依托单位:
Developmental reinforcement learning
-
批准号:238988-2010
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$4.37万
-
财政年份:2010
-
负责人:Precup, Doina
-
依托单位:
Learning and prediction in high-dimensional stochastic environments
-
批准号:335248-2005
-
项目类别:Collaborative Research and Development Grants
-
资助金额:$5.5万
-
财政年份:2009
-
负责人:Precup, Doina
-
依托单位:
Knowledge representation in reinforcement learning
-
批准号:238988-2005
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$2.19万
-
财政年份:2009
-
负责人:Precup, Doina
-
依托单位:
Knowledge representation in reinforcement learning
-
批准号:238988-2005
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$2.19万
-
财政年份:2008
-
负责人:Precup, Doina
-
依托单位:
海外基金