NRI: INT: SCHooL: Scalable Collaborative Human-Robot Learning
NRI: INT: SCHooL: Scalable Collaborative Human-Robot Learning
批准号:
1734633
负责人:
Ken Goldberg
金额:
$137.49万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2017
资助国家:
美国
项目状态:
已结题
起止时间:
2017-09-01 至 2021-08-31
中文摘要
为了在仓库、家庭和从学校到零售店的其他环境中发挥作用,机器人需要学习如何稳健地操纵各种各样的物体。 例如,为了提高人类工人的生产力,服务和工厂机器人可以通过识别、抓取和将物体重新定位到适当的位置来保持指定表面的清洁。 对机器人进行预编程以执行如此复杂的操作任务是不可行的;相反,该项目将研究可扩展的机器人操作,其中多个机器人协同向多个人类学习。 该项目将贡献新的模型、算法、软件和实验数据,以推进深度学习、人机交互和云机器人技术的发展。 为了向学生和公众广泛传达这项研究的结果,该项目将与劳伦斯科学馆和非洲机器人网络合作制作一本书和视频。目前对合作机器人演示学习(LfD)的理解存在两个主要差距:1)缺乏一个包含人类和机器人的理论框架,以产生合作学习行为作为最佳解决方案; 2)缺乏将LfD与深度学习,分层规划和人机交互联系起来的研究。该项目通过基于逆向强化学习和人类与机器人之间通信的博弈论模型的统一理论框架来解决这些差距,将LfD视为一个可扩展的协作机器人过程,其中多个人类和多个联网机器人在一组分布式环境中工作,以最大化集体奖励函数,人类学习如何成为机器人更有效的演示者。这项研究几乎可以应用于任何机器人可以从人类演示中学习的环境,并将在项目过程中以越来越复杂的“表面清理”基准进行评估。
英文摘要
To be useful in warehouses, homes, and other environments from schools to retail stores, robots will need to learn how to robustly manipulate a wide variety of objects. For instance, to enhance the productivity of human workers, service and factory robots could keep specified surfaces clear by identifying, grasping, and relocating objects to appropriate locations. Pre-programming robots to perform such complex manipulation tasks is not feasible; instead this project will investigate scalable robot manipulation, where multiple robots collaboratively learn from multiple humans. The project will contribute new models, algorithms, software, and experimental data to advance the state-of-the-art in deep learning, human-robot interaction, and cloud robotics. To broadly convey the results of this research to students and the public, the project will create a book and video with the Lawrence Hall of Science and the African Robotics Network. Two primary gaps in current understanding of co-robotic Learning from Demonstration (LfD) are: 1) the absence of a theoretical framework that encompasses humans and robots to produce cooperative learning behaviors as optimal solutions; and 2) the lack of research linking LfD with deep learning, hierarchical planning, and human-robot interaction. The project addresses those gaps with a unified theoretical framework based on Inverse Reinforcement Learning and game-theoretic models of communication between humans and robots, treating LfD as a scalable co-robotic process in which multiple humans and multiple networked robots work in a distributed set of environments to maximize a collective set of reward functions and humans learn how to become more effective demonstrators for robots. The research can be applied to almost any context where robots can learn from human demonstrations and will be evaluated in "surface decluttering" benchmarks of increasing complexity over the course of the project.
期刊论文(59)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
DOI:
--
发表时间:
2017-03
期刊:
影响因子:
--
作者:
[Michael Laskey;Jonathan Lee;Roy Fox;A. Dragan;Ken Goldberg]
通讯作者:
Michael Laskey;Jonathan Lee;Roy Fox;A. Dragan;Ken Goldberg
DOI:
10.1109/lra.2020.2976272
发表时间:
2019-05
期刊:
IEEE Robotics and Automation Letters
影响因子:
5.2
作者:
[Brijen Thananjeyan;A. Balakrishna;Ugo Rosolia;Felix Li;R. McAllister;Joseph Gonzalez;S. Levine;F. Borrelli;Ken Goldberg]
通讯作者:
Brijen Thananjeyan;A. Balakrishna;Ugo Rosolia;Felix Li;R. McAllister;Joseph Gonzalez;S. Levine;F. Borrelli;Ken Goldberg
DOI:
--
发表时间:
2019-12
期刊:
影响因子:
--
作者:
[S. Reddy;A. Dragan;S. Levine;S. Legg;J. Leike]
通讯作者:
S. Reddy;A. Dragan;S. Levine;S. Legg;J. Leike
DOI:
10.1007/s10514-018-9771-0
发表时间:
2019-02-01
期刊:
AUTONOMOUS ROBOTS
影响因子:
3.5
作者:
[Huang, Sandy H., Held, David, Dragan, Anca D.]
通讯作者:
Dragan, Anca D.
DOI:
10.24963/ijcai.2017/196
发表时间:
2017-08
期刊:
影响因子:
--
作者:
[Aijun Bai;Stuart J. Russell]
通讯作者:
Aijun Bai;Stuart J. Russell
共 55 条
CONE: Collaborative Observatory for Natural Environments
-
批准号:0535218
-
项目类别:Continuing Grant
-
资助金额:$0.0万
-
财政年份:2005
-
负责人:Ken Goldberg
-
依托单位:
Self-Aligning Grippers and Fixtures
-
批准号:0010069
-
项目类别:Standard Grant
-
资助金额:$10.0万
-
财政年份:2001
-
负责人:Ken Goldberg
-
依托单位:
ITR/SI: Collaborative Telerobotics: Theory and Scalable Infrastructure
-
批准号:0113147
-
项目类别:Continuing Grant
-
资助金额:$45.0万
-
财政年份:2001
-
负责人:Ken Goldberg
-
依托单位:
Postdoc: ECS Postdoctoral Associate: Design and Simulation of Algorithms and Mechanisms for Precision Assembly
-
批准号:9705022
-
项目类别:Standard Grant
-
资助金额:$4.61万
-
财政年份:1997
-
负责人:Ken Goldberg
-
依托单位:
Challenges in CISE: Planning and Control for Massively Parallel Manipulation
-
批准号:9726389
-
项目类别:Continuing Grant
-
资助金额:$136.62万
-
财政年份:1997
-
负责人:Ken Goldberg
-
依托单位:
Reduced Complexity Manipulation with the Parallel-Jaw Gripper: 2-Year Plan for Extended Research
-
批准号:9612491
-
项目类别:Continuing Grant
-
资助金额:$12.0万
-
财政年份:1996
-
负责人:Ken Goldberg
-
依托单位:
Presidential Faculty Fellowships
-
批准号:9553197
-
项目类别:Continuing Grant
-
资助金额:$37.5万
-
财政年份:1996
-
负责人:Ken Goldberg
-
依托单位:
NSF Young Investigator
-
批准号:9696061
-
项目类别:Continuing Grant
-
资助金额:$26.44万
-
财政年份:1995
-
负责人:Ken Goldberg
-
依托单位:
NSF Young Investigator
-
批准号:9457523
-
项目类别:Continuing Grant
-
资助金额:$8.75万
-
财政年份:1994
-
负责人:Ken Goldberg
-
依托单位:
Reduced Complexity Manipulation with the Parallel-Jaw Gripper
-
批准号:9123747
-
项目类别:Continuing Grant
-
资助金额:$31.5万
-
财政年份:1992
-
负责人:Ken Goldberg
-
依托单位:
国内基金
海外基金
登录
查看更多内容
内源性逆转录病毒MER65-int调控人类胎
盘发育与子宫内膜重塑的功能研究
-
批准号:
-
项目类别:省市级项目
-
资助金额:10.0万元
-
批准年份:2025
-
负责人:屈雨亮
-
依托单位:
隐秘重组信号序列INT-RSS在T细胞受体基因Tcra重排中的功能和机制研究
-
批准号:32370939
-
项目类别:面上项目
-
资助金额:50万元
-
批准年份:2023
-
负责人:郝冰涛
-
依托单位:
HPV16 E7 通过 Int1 蛋白调控 Wnt 信号通路调节肿瘤局部树突状细胞活性
-
批准号:LQ22H160033
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2021
-
负责人:陈婷婷
-
依托单位:
选择性PPARγ激动剂INT131调控适应性产热和AD-MSCs分化成棕色样脂肪细胞的机制研究
-
批准号:81903680
-
项目类别:青年科学基金项目
-
资助金额:20.0万元
-
批准年份:2019
-
负责人:高茸
-
依托单位:
INT复合物调节U snRNA 3'加工的结构基础
-
批准号:31800624
-
项目类别:青年科学基金项目
-
资助金额:28.0万元
-
批准年份:2018
-
负责人:杭婧
-
依托单位:
沉默Int6基因的骨髓间充质干细胞复合生物支架构建血管化腹股沟疝补片及其促补片血管化机制
-
批准号:81371698
-
项目类别:面上项目
-
资助金额:70.0万元
-
批准年份:2013
-
负责人:赵一麟
-
依托单位:
HIF/Int6调控迟发型EPC体外增殖的机制及其治疗重度子痫前期的可行性
-
批准号:81100439
-
项目类别:青年科学基金项目
-
资助金额:22.0万元
-
批准年份:2011
-
负责人:李勤
-
依托单位: