Active Reward Learning from Online Preferences

Active Reward Learning from Online Preferences
复制标题

根据在线偏好进行主动奖励学习

DOI:
--
复制
发表时间:
2023
期刊:
International Conference on Robotics and Automation (ICRA
影响因子:
--
通讯作者:
Sadigh, Dorsa
Sadigh, Dorsa
中科院分区:
--
文献类型:
--
作者:
Myers, Vivek;Biyik, Erdem;Sadigh, Dorsa

文献摘要

参考文献

被引文献

相似文献

根据专家偏好学习城市空中交通遭遇模型
DOI: 10.1109/dasc43569.2019.9081648
发表时间: 2019
期刊: 2019 IEEE/AIAA 38th Digital Avionics Systems Conference (DASC)
影响因子: --
作者:
Sydney M. Katz;Anne;Mykel J. Kochenderfer
通讯作者: Mykel J. Kochenderfer
DOI: 10.1109/iros.2011.6094735
发表时间: 2011-12
期刊: 2011 IEEE/RSJ International Conference on Intelligent Robots and Systems
影响因子: --
作者:
M. Cakmak;S. Srinivasa;Min Kyung Lee;J. Forlizzi;S. Kiesler
通讯作者: M. Cakmak;S. Srinivasa;Min Kyung Lee;J. Forlizzi;S. Kiesler
这是我学到的:提出能够奖励学习的问题
DOI: --
发表时间: 2021
影响因子: 5.1
作者:
Soheil Habibian;Ananth Jonnavittula;Dylan P. Losey
通讯作者: Dylan P. Losey
为深度强化学习代理提供不确定性感知行动建议
DOI: --
发表时间: 2020
期刊: AAAI Conference on Artificial Intelligence
影响因子: --
作者:
Felipe Leno da Silva;Pablo Hernandez;Bilal Kartal;Matthew Taylor
通讯作者: Matthew Taylor
DOI: --
发表时间: 2019-04
期刊: --
影响因子: --
作者:
Daniel S. Brown;Wonjoon Goo;P. Nagarajan;S. Niekum
通讯作者: Daniel S. Brown;Wonjoon Goo;P. Nagarajan;S. Niekum