Reframing Deep Reinforcement Learning for Large-Scale, Real-World Implementation
Reframing Deep Reinforcement Learning for Large-Scale, Real-World Implementation
批准号:
2373874
负责人:
金额:
$0.0万
依托单位:
依托单位国家:
英国
项目类别:
Studentship
财政年份:
2019
资助国家:
英国
项目状态:
已结题
起止时间:
2019 至 --
中文摘要
应用于机器人系统的强化学习有可能彻底改变许多行业,但仍然缺乏部署所需的灵活性,可扩展性和安全性。这种方法的一个关键瓶颈是依赖于从手工设计的奖励函数中获取信息,需要在经验昂贵的任务中进行非平凡的设计工作,本研究项目的目标是增强经典的强化学习方法,以开发高效,实用和可扩展的机器人控制算法。最终目标是生产和部署一个多功能和安全的系统,能够学习一组可定制的任务,从工业流程到运行时的人工支持。这需要提供自主代理的能力,以恢复其环境的结构化表示,获得层次知识的更高层次的行为,并利用社会信号,以最好地了解他们的目标。目前的计划是单独实现这些子目标,并最终在丰田HSR机器人上共同实现。预期的研究成果是:-开发一种新的方法,使人类互动的使用,形成一个有意义的形式的信息学习新的任务。定义一个框架,用于无监督地恢复代理机械输入的更高级别的解纠缠表示。整合元学习系统,在一系列用户指定的任务中提高培训效率。将这些系统移植到HSR机器人上,并评估与人类用户交互的真实场景。
英文摘要
Reinforcement learning applied to robotics systems has the potential to revolutionize many industries, but still lacks the flexibility, scalability and safety needed for deployment. One key bottleneck of such methods is their reliance on obtaining information from hand-engineered reward functions, requiring a non-trivial design effort in tasks where experience is expensive to collect.The goal of this research project is to augment the classical reinforcement learning approach to develop efficient, practical and scalable algorithms for robotics control. The ultimate objective is to produce and deploy a versatile and safe system, able to learn a set of customizable tasks ranging from industrial processes to human support at run time. This entails providing autonomous agents with the ability to recover a structured representation of their environment, obtain hierarchical knowledge over higher-level behaviour and utilize social signals to best understand their objectives. The current plan is to approach each of these sub-goals individually and ultimately implement them jointly on the Toyota HSR Robot. The expected research outcomes are:- Develop a novel method to enable usage of human interaction to form a meaningful form of information for learning new tasks.- Define a framework for the unsupervised recovery of a higher-level disentangled representation of the agent's mechanical inputs.- Incorporate a system for meta-learning, bolstering training efficiency over a range of user-specified tasks.- Port these systems on the HSR robot and evaluate real-world scenarios of interaction with human users.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
国内基金
海外基金
登录
查看更多内容
Deep Seek引导下预防肝硬化腹水患者发生腹腔感染的约翰霍普金斯循证实践模型下中医护理策略的构建研究
-
批准号:2026JJ81909
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2026
-
负责人:胡曦
-
依托单位:
基于Deep Unrolling的高分辨近红外二区荧光分子断层成像方法研究
-
批准号:12271434
-
项目类别:面上项目
-
资助金额:46万元
-
批准年份:2022
-
负责人:贺小伟
-
依托单位:
基于深度森林(Deep Forest)模型的表面增强拉曼光谱分析方法研究
-
批准号:2020A151501709
-
项目类别:省市级项目
-
资助金额:10.0万元
-
批准年份:2020
-
负责人:谢怡
-
依托单位:
面向Deep Web的数据整合关键技术研究
-
批准号:61872168
-
项目类别:面上项目
-
资助金额:62.0万元
-
批准年份:2018
-
负责人:董永权
-
依托单位:
基于Deep-learning的三江源区冰川监测动态识别技术研究
-
批准号:51769027
-
项目类别:地区科学基金项目
-
资助金额:38.0万元
-
批准年份:2017
-
负责人:张大奇
-
依托单位:
具有时序处理能力的Spiking-Deep Learning(脉冲深度学习)方法研究
-
批准号:61573081
-
项目类别:面上项目
-
资助金额:64.0万元
-
批准年份:2015
-
负责人:屈鸿
-
依托单位:
基于语义计算的海量Deep Web知识探索机制研究
-
批准号:61272411
-
项目类别:面上项目
-
资助金额:80.0万元
-
批准年份:2012
-
负责人:赵峰
-
依托单位:
Deep Web数据集成查询结果抽取与整合关键技术研究
-
批准号:61100167
-
项目类别:青年科学基金项目
-
资助金额:20.0万元
-
批准年份:2011
-
负责人:董永权
-
依托单位:
面向Deep Web的大规模知识库自动构建方法研究
-
批准号:61170020
-
项目类别:面上项目
-
资助金额:57.0万元
-
批准年份:2011
-
负责人:崔志明
-
依托单位:
Deep Web敏感聚合信息保护方法研究
-
批准号:61003054
-
项目类别:青年科学基金项目
-
资助金额:20.0万元
-
批准年份:2010
-
负责人:赵朋朋
-
依托单位: