CRII: RI: Accelerated Stochastic Approximation for Reinforcement Learning
CRII: RI: Accelerated Stochastic Approximation for Reinforcement Learning
批准号:
1566186
负责人:
Martha White
金额:
$17.46万
依托单位:
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2016
资助国家:
美国
项目状态:
已结题
起止时间:
2016-06-01 至 2018-12-31
中文摘要
点击翻译按钮获取中文摘要
英文摘要
This project develops a new class of accelerated learning techniques for reinforcement learning. Reinforcement learning is an approach to autonomous decision-making through trial-and-error interaction with an unknown environment, with a focus on learning incrementally from this stream of data. Reinforcement learning has significant industrial potential, particularly for real-time control systems, such as active network management for energy and search-and-rescue robots, and is already used in a wide range of fields, including robotics, psychology, animal learning and neuroscience. To improve the practical application of reinforcement learning, this project proposes a new class of algorithms with the goal to balance computational complexity and the sample efficiency of learning, which often requires significant computation and memory. This space of algorithms that attempt to balance both requirements has been under-explored for reinforcement learning, and provide exciting opportunities to impact industrial applications and the growing area of computational sustainability. An important aspect of this project will be to implement and study these algorithms on a wide-range of simulated environments, and engage a diverse group of students through courses and summer research.This project develops efficient incremental approximations to summarize gathered samples for improved sample efficiency and an empirical framework to evaluate these algorithms. This new class of accelerated learning techniques formally trade-off computation and accuracy and have many promising extensions and research directions, through a variety of accelerated stochastic gradient descent techniques and incremental matrix approximations. Further, another focus is to develop tools and novel measures for the reinforcement learning community that evaluate this balance between sample efficiency and computational complexity, with the code framework released through an existing open-source platform. This initial systematic exploration of these novel optimization variants will lay the foundation for the long-term goal of improving efficacy of reinforcement learning in industry and for practical autonomous agents.
期刊论文(1)
专著(0)
科研奖励(0)
会议论文
DOI:
10.1609/aaai.v33i01.33014384
发表时间:
2018-11
期刊:
影响因子:
--
作者:
[Vincent Liu;Raksha Kumaraswamy;Lei Le;Martha White]
通讯作者:
Vincent Liu;Raksha Kumaraswamy;Lei Le;Martha White
国内基金
海外基金
登录
查看更多内容
破骨细胞源性FcγRI介导类风湿性关节炎炎症后疼痛的作用机制
-
批准号:
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2026
-
负责人:阳林
-
依托单位:
四神丸调控生物钟基因Bmal1/Fc εRI介导肥大细胞节律性活化治疗IBS-D“晨起痛”的作用机制研究
-
批准号:
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2026
-
负责人:何心凌
-
依托单位:
NSUN6介导的m5C修饰调控心肌细胞凋亡和铁死亡参与MI/RI的机制研究
-
批准号:2026JJ80739
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2026
-
负责人:袁乐宏
-
依托单位:
中药牛耳枫中抗MI/RI新颖虎皮楠生物碱的发现与作用机制研究
-
批准号:2026JJ60255
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2026
-
负责人:张济辉
-
依托单位:
醒脑静多靶点调控PI3K/Akt通路抑制CI/RI氧化应激—基于网络药理学及体内、外实验研究
-
批准号:2025JJ90117
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2025
-
负责人:李秋云
-
依托单位:
IgA-FcαRI介导的Syk/NLRP3/caspase-1通路在线状IgA大疱性皮病
中的机制研究
-
批准号:
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2024
-
负责人:荆可
-
依托单位:
基于双修饰ANG-RNH1系统阻抑RI复合物生成机制建立口腔黏膜等效物血管化稳态
-
批准号:82401112
-
项目类别:青年科学基金项目
-
资助金额:30万元
-
批准年份:2024
-
负责人:刘旭倩
-
依托单位:
跨膜蛋白LRP5胞外域调控膜受体TβRI促钛表面BMSCs归巢、分化的研究
-
批准号:82301120
-
项目类别:青年科学基金项目
-
资助金额:30万元
-
批准年份:2023
-
负责人:於科
-
依托单位:
基于“免疫-神经”网络探讨眼针活化CI/RI大鼠MC靶向H3R调节“免疫监视”的抗炎机制
-
批准号:82374375
-
项目类别:面上项目
-
资助金额:51万元
-
批准年份:2023
-
负责人:马贤德
-
依托单位:
Dectin-2通过促进FcεRI聚集和肥大细胞活化加剧哮喘发作的机制研究
-
批准号:82300022
-
项目类别:青年科学基金项目
-
资助金额:30万元
-
批准年份:2023
-
负责人:屈玉兰
-
依托单位: